Large Language Models

Explore language models, how they work, and the techniques used to build applications with them.

Explore articles

Reset filters
Browse subtopics: AI Models

Articles that also belong to these categories. Counts cover all of Large Language Models.

Showing 121-180 of 184 articles

Ling-3.0-flash

Ling-3.0-flash is an open-weight mixture-of-experts language model from inclusionAI, the open-source AI initiative of Ant Group, and the first model of the Ling 3.0 generation.

AI ModelsChinese AI

Llama 3

Llama 3 is a family of open-weight large language models developed by Meta. Meta released the original Llama 3 checkpoints on April 18, 2024, in 8-billion-parameter and 70-billion-parameter sizes.

AI ModelsMeta AI

Llama 3.1

Llama 3.1 is a family of open-weight large language models released by Meta on July 23, 2024, in three sizes, 8 billion, 70 billion, and 405 billion parameters, each shipped in both a pre-trained base form and…

AI ModelsMeta AI

Llama 3.2

Llama 3.2 is a family of four open-weight large language models released by Meta on September 25, 2024, comprising lightweight 1 billion and 3 billion parameter text-only models for on-device AI and the 11…

AI ModelsMeta AI

Llama 3.3

Llama 3.3 is an instruction-tuned, text-only large language model with 70 billion parameters that Meta released on December 6, 2024

AI ModelsMeta AI

LongCat-Flash

LongCat-Flash is an open-weight large language model developed by the LongCat team at Meituan, the Chinese on-demand local-services and food-delivery company.

AI ModelsOpen Source AI

MAI-1-preview

MAI-1-preview is a large language model developed by Microsoft AI, the consumer artificial intelligence division of Microsoft led by Mustafa Suleyman.

AI ModelsOpen Source AI

MiniCPM

MiniCPM is a family of compact, openly licensed language models published by OpenBMB, the shared open-source brand of Tsinghua University's natural language processing lab (THUNLP) and the Beijing company…

AI ModelsChinese AI

MiniCPM5-2B

MiniCPM5-2B is an open-weights small language model published by OpenBMB on September 7, 2026 under the Apache License 2.0.

AI ModelsChinese AI

MiniMax M2

MiniMax M2 is an open-weight large language model released on October 27, 2025 by the Shanghai-based AI company MiniMax, built as a Mixture of Experts model with 230 billion total parameters and roughly 10…

AI AgentsAI Models

Mistral Large

Mistral Large is the family of flagship large language models developed by Mistral AI, the Paris-based AI laboratory founded in 2023, and is the company's most capable general-purpose model line.

AI CompaniesAI Models

Mistral Large 3

Mistral Large 3 is a sparse mixture-of-experts large language model released on December 2, 2025 by the French AI company Mistral AI, distributed as open weights under the Apache 2.0 license with roughly 675…

AI CompaniesAI Models

Muse Glimmer

Muse Glimmer is an open-weight text-and-image model developed by Meta AI for local agent and coding workloads. Meta released the model on August 10, 2026 under the identifier meta-models/Muse-Glimmer-30B.

AI AgentsAI Models

Nemotron

Nemotron is NVIDIA's brand for its family of open large language models and the datasets, training recipes, and evaluation tools built around them.

AI ModelsNVIDIA

Nemotron 3

Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…

AI ModelsNVIDIA

Nemotron-4

Nemotron-4 is a family of decoder-only large language models developed by NVIDIA and documented in two technical reports released in 2024.

AI ModelsNVIDIA

Nemotron-H

Nemotron-H is a family of open-weight large language models released by NVIDIA in April 2025 that replace most of the self-attention layers of a standard Transformer with Mamba-2 state-space layers, producing…

AI ModelsNVIDIA

OLMo

OLMo (Open Language Model) is a family of fully open large language models built by the Allen Institute for AI (Ai2) and first released on February 1, 2024.

AI ModelsOpen Source AI

OLMo 2

OLMo 2 is the second generation of fully open large language models released by the Allen Institute for AI (Ai2), spanning 7B, 13B, and 32B parameter sizes.

AI ModelsOpen Source AI

OLMoE

OLMoE (Open Mixture-of-Experts) is a fully open sparse mixture of experts large language model released by the Allen Institute for AI (Ai2) on September 3, 2024 .

AI ModelsMixture of Experts

OpenAI o1

OpenAI o1 is a family of proprietary large language models developed by OpenAI and trained to use additional computation before returning an answer.

AI ModelsOpenAI

OpenAI o3

OpenAI o3 is a family of reasoning-focused large language models developed by OpenAI and the second generation of the company's o-series reasoning models, best known for scoring 87.5% on the ARC-AGI…

AI ModelsOpenAI

Phi-4

Phi-4 is a 14-billion-parameter small language model developed by Microsoft Research and released in December 2024, designed to match or beat models several times its size on reasoning tasks by training…

AI ModelsMicrosoft

Phi-4-mini

Phi-4-mini is a 3.8 billion parameter open weight small language model released by Microsoft on February 26, 2025, under the permissive MIT license.

AI ModelsOpen Source AI

Pixtral

Pixtral is a family of multimodal vision-language models developed by Mistral AI, a French AI company founded in April 2023.

AI CompaniesAI Models

Qwen3

Qwen3 is the third-generation family of large language models developed by the Qwen Team at Alibaba Cloud (also known as Tongyi Qianwen lab).

AI ModelsChinese AI

Qwen3-Max

Qwen3-Max is the flagship large language model in Alibaba's Qwen series and the first Qwen model to cross one trillion parameters, released in preview on September 5, 2025 and formally launched at the Apsara…

AI ModelsChinese AI

Qwen3-Next

Qwen3-Next is an efficiency-focused large language model and model architecture released in September 2025 by the Qwen team at Alibaba Cloud.

AI ModelsOpen Source AI

Qwen3.8

Qwen3.8 is the name Qwen uses for a model generation that includes hosted services and downloadable checkpoints.

AI ModelsChinese AI

Reka Core

Reka Core is a frontier class multimodal foundation model developed by Reka AI, a research and product company founded in 2022 by former scientists from DeepMind, Google Brain, Meta FAIR, and Baidu.

AI ModelsMultimodal AI

Reka Edge

Reka Edge is a 7-billion-parameter multimodal language model developed by Reka AI, introduced in April 2024 as the smallest member of the company's first publicly described model family.

AI ModelsMultimodal AI

Reka Flash

Reka Flash is a family of multimodal large language models developed by Reka AI, a San Francisco Bay Area research company founded in 2022 by former researchers from Google DeepMind, Meta FAIR, and Google.

AI ModelsMultimodal AI

Ring-1T

Ring-1T is an open-weight reasoning model released in October 2025 by inclusionAI, the open-source research initiative associated with Ant Group

AI ModelsOpen Source AI

Skywork R1V

Skywork R1V is a family of open-weight multimodal vision-language models built for chain-of-thought reasoning, developed by Skywork AI

AI Models

SmolLM

SmolLM is a family of small, fully open language models released by Hugging Face on July 16, 2024 in three sizes, 135 million, 360 million, and 1.7 billion parameters, all trained on a curated open dataset…

AI ModelsOpen Source AI

SmolLM 3

SmolLM 3 is a fully open 3 billion parameter language model released by Hugging Face on July 8, 2025, trained on 11.2 trillion tokens and designed as a small, multilingual, long-context reasoner.

AI ModelsOpen Source AI

Step-3

Step-3 is an open-weight large multimodal mixture of experts (MoE) model released in July 2025 by StepFun, the Shanghai-based Chinese artificial intelligence startup also known as Jieyue Xingchen.

AI Models

Yi-Large

Yi-Large is a closed-source large language model developed by Chinese artificial intelligence company 01.AI (零一万物, Língyi Wànwù), founded by Kai-Fu Lee.

AI ModelsChinese AI

Yi-Lightning

Yi-Lightning is a closed-source large language model developed by Chinese artificial intelligence company 01.AI (零一万物, Língyī Wànwù), the company founded by Kai-Fu Lee.

AI ModelsChinese AI