Large Language Models

Explore language models, how they work, and the techniques used to build applications with them.

Explore articles

Reset filters
Browse subtopics: NVIDIA

Articles that also belong to these categories. Counts cover all of Large Language Models.

Showing 1-13 of 13 articles

ChipNeMo

ChipNeMo is a research project and a family of domain-adapted large language models developed by Nvidia to assist with industrial semiconductor and chip-design tasks.

AI HardwareNVIDIA

Llama Nemotron

Llama Nemotron is a family of open reasoning large language models built by Nvidia by post-training Meta's Llama models for math, coding, and agentic tasks.

NVIDIAReasoning Models

Mistral NeMo

Mistral NeMo is a 12 billion parameter large language model released by Mistral AI in collaboration with NVIDIA on July 18, 2024 .

NVIDIAOpen Source AI

NVLM

NVLM (short for NVIDIA Vision Language Model), released as NVLM 1.0, is a family of open multimodal large language models developed by Nvidia.

Multimodal AINVIDIA

Nemotron

Nemotron is NVIDIA's brand for its family of open large language models and the datasets, training recipes, and evaluation tools built around them.

AI ModelsNVIDIA

Nemotron 3

Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…

AI ModelsNVIDIA

Nemotron-4

Nemotron-4 is a family of decoder-only large language models developed by NVIDIA and documented in two technical reports released in 2024.

AI ModelsNVIDIA

Nemotron-H

Nemotron-H is a family of open-weight large language models released by NVIDIA in April 2025 that replace most of the self-attention layers of a standard Transformer with Mamba-2 state-space layers, producing…

AI ModelsNVIDIA

SparDA

SparDA (Sparse Decoupled Attention) is an add-on architecture for long-context large language model inference proposed by researchers at NVIDIA in a paper posted to arXiv on 3 June 2026.

AI InferenceModel Architecture