Jet-Nemotron
Jet-Nemotron is a family of small hybrid-architecture language models released by NVIDIA Research in August 2025.
Explore NVIDIA through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of NVIDIA.
Showing 1-28 of 28 articles
Jet-Nemotron is a family of small hybrid-architecture language models released by NVIDIA Research in August 2025.
KAI Scheduler is an open-source Kubernetes scheduler that optimizes the allocation of GPU resources for artificial intelligence and machine learning workloads.
Llama-3.1-Nemotron-70B-Instruct is a large language model released by NVIDIA in October 2024.
Megatron-LM is NVIDIA's open-source framework for training very large transformer language models across GPU clusters
Mistral NeMo is a 12 billion parameter large language model released by Mistral AI in collaboration with NVIDIA on July 18, 2024 .
NOOA (NVIDIA Object-Oriented Agents) is an open-source, model-agnostic Python framework for building AI agents, released by NVIDIA in July 2026 .
NVIDIA BioNeMo is NVIDIA's software platform for applying AI to biology and drug discovery.
NVIDIA BioNeMo Inference Runtime (BioIR) is a GPU-accelerated Python library for biomolecular structure prediction inference.
NVIDIA Dynamo is an open-source, low-latency distributed inference serving framework designed to deploy and scale generative AI and reasoning models across large GPU clusters.
NVIDIA Isaac Lab-Arena is an open-source framework for composing simulated robot tasks and evaluating learned policies at scale.
NVIDIA Ising is a family of open AI models from NVIDIA for operating quantum computers, covering two of the field's main engineering bottlenecks: quantum processor calibration and quantum error-correction…
The NVIDIA NeMo Agent Toolkit is an open-source, framework-agnostic library for connecting, profiling, evaluating, and optimizing teams of AI agents.
NVIDIA NeMo Switchyard is an open-source proxy and Rust library for routing requests among configured large language models.
NVIDIA NemoClaw is a collection of open blueprints (reference architectures) for building custom autonomous AI agents that packages three moving parts, a model, an agent harness, and a secure runtime
Newton is an open-source, GPU-accelerated physics engine built for robotics simulation and robot learning, co-developed by NVIDIA, Google DeepMind, and Disney Research and stewarded by the Linux Foundation.
NVIDIA OSMO is an open-source workflow orchestration platform developed by Nvidia for physical AI and robotics development.
NVIDIA OpenShell is an open source runtime that executes autonomous AI agents inside policy-governed sandboxes, published by NVIDIA under the Apache License 2.0.
Parakeet is a family of open automatic speech recognition (ASR) models developed by NVIDIA as part of the NeMo conversational AI toolkit.
NVIDIA TensorRT-LLM is an open-source library developed by nvidia for high-performance inference of large language models on NVIDIA GPUs.
NVIDIA Warp is an open-source Python framework, first released by NVIDIA in March 2022, that takes ordinary Python functions and just-in-time compiles them into native kernels that run on the CPU or on a…
The NVIDIA acquisition of Hugging Face is a pending transaction in which NVIDIA agreed to buy Hugging Face, the New York company that operates the largest hosting platform for open machine-learning models and…
Nemotron is NVIDIA's brand for its family of open large language models and the datasets, training recipes, and evaluation tools built around them.
Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…
NVIDIA Nemotron 3.5 Lightning is an open-weights 30 billion parameter mixture-of-experts language model with 3 billion active parameters per token, released by NVIDIA on August 11
Nemotron-4 is a family of decoder-only large language models developed by NVIDIA and documented in two technical reports released in 2024.
Nemotron-H is a family of open-weight large language models released by NVIDIA in April 2025 that replace most of the self-attention layers of a standard Transformer with Mamba-2 state-space layers, producing…
Nemotron-Labs-TwoTower is an open-weight diffusion language model released by NVIDIA in mid-2026.
RAPIDS is an open-source suite of GPU-accelerated software libraries for data science, analytics, and machine learning, developed and maintained by Nvidia.