Large Language Models

Explore language models, how they work, and the techniques used to build applications with them.

Explore articles

Reset filters
Browse subtopics: Reasoning Models

Articles that also belong to these categories. Counts cover all of Large Language Models.

Showing 1-39 of 39 articles

Claude Opus 5

Claude Opus 5 is a large language model released by Anthropic on 24 July 2026, the newest entry in the Claude Opus line and the successor to Claude Opus 4.8.

AI ModelsAnthropic

DeepSeek V3.1

DeepSeek V3.1 is a large language model developed by DeepSeek, released on August 19, 2025 and made broadly available via the official API on August 21, 2025.

AI ModelsChinese AI

DeepSeek-Prover

DeepSeek-Prover is a family of open-weight large language models developed by Chinese AI laboratory DeepSeek for formal theorem proving in the Lean 4 proof assistant.

Chinese AIReasoning Models

DeepSeek-R1

DeepSeek-R1 is an open-weight reasoning model and large language model family developed by the Chinese artificial intelligence laboratory DeepSeek. The original model was released on January 20, 2025.

Chinese AIReasoning Models

DeepSeek-R1-Distill

DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.

AI ModelsChinese AI

DeepSeekMath

DeepSeekMath is a family of open-weight large language models specialized for mathematical reasoning, released by Chinese AI laboratory DeepSeek in February 2024.

Chinese AIReasoning Models

GPT-5.6

GPT-5.6 is a family of proprietary multimodal large language models developed by OpenAI. The family entered a limited preview on June 26, 2026, and became generally available on July 9, 2026.

AI ModelsOpenAI

GPT-6 Astra

GPT-6 Astra is an OpenAI model that the company began rolling out on September 3, 2026, describing it as "the world's most intelligent and aligned model" and as the successor to the GPT-5.6 family.

AI ModelsOpenAI

GSM8K

GSM8K (Grade School Math 8K) is an English-language benchmark of grade-school arithmetic word problems released by OpenAI researchers in 2021.

AI BenchmarksMachine Learning

Grok 4

Grok 4 is a large language model developed by xAI and released on July 9, 2025. It is the fourth major generation of the Grok model family and was positioned as xAI's most capable model to date at its release.

AI CompaniesAI Models

Grok 4.5

Grok 4.5 is a proprietary multimodal large language model and reasoning model in the Grok family. It was developed by SpaceXAI in collaboration with Cursor and released through the xAI API on July 8, 2026.

AI ModelsGenerative AI

Grok 4.6

Grok 4.6 is a proprietary large language model and reasoning model in the Grok family, developed by SpaceXAI and released jointly with Cursor through the xAI API on August 12, 2026.

AI ModelsGenerative AI

Llama Nemotron

Llama Nemotron is a family of open reasoning large language models built by Nvidia by post-training Meta's Llama models for math, coding, and agentic tasks.

NVIDIAReasoning Models

MiniMax M1

MiniMax M1 (stylised MiniMax-M1) is an open-weight large language reasoning model released on 16 June 2025 by the Shanghai-based artificial-intelligence company MiniMax

Chinese AIReasoning Models

OpenAI o-series

The OpenAI o-series is a family of large language models developed by OpenAI that are trained with reinforcement learning to reason through an internal chain-of-thought before answering, making them OpenAI's…

Artificial IntelligenceOpenAI

OpenAI o1

OpenAI o1 is a family of proprietary large language models developed by OpenAI and trained to use additional computation before returning an answer.

AI ModelsOpenAI

OpenAI o1-mini

OpenAI o1-mini is a smaller, faster, and cheaper reasoning model released by OpenAI on September 12, 2024, alongside o1-preview, and optimized for science, technology, engineering, and mathematics (STEM) tasks…

OpenAIReasoning Models

OpenAI o1-pro

OpenAI o1-pro is the highest-compute variant of OpenAI's o1 reasoning model, designed to spend more inference-time compute so it "thinks harder" and returns the most reliable answers on the hardest…

OpenAIReasoning Models

OpenAI o3

OpenAI o3 is a family of reasoning-focused large language models developed by OpenAI and the second generation of the company's o-series reasoning models, best known for scoring 87.5% on the ARC-AGI…

AI ModelsOpenAI

OpenAI o3-mini

OpenAI o3-mini is a reasoning-focused large language model released by OpenAI on January 31, 2025, the second commercial member of the o-series after OpenAI o1 and a smaller, cheaper

OpenAIReasoning Models

OpenAI o3-pro

OpenAI o3-pro is a high-compute reasoning large language model released by OpenAI on June 10, 2025, designed as the professional, higher-reliability variant of the company's o3 reasoning model.

OpenAIReasoning Models

QwQ

QwQ is a family of open-weight reasoning models from the Qwen team at Alibaba Cloud, built to compete with OpenAI's o1 and DeepSeek-R1 at a fraction of their size.

Chinese AIOpen Source AI

Qwen3.8

Qwen3.8 is the name Qwen uses for a model generation that includes hosted services and downloadable checkpoints.

AI ModelsChinese AI

Self-consistency

Self-consistency is a decoding strategy for large language models that samples multiple chain-of-thought reasoning paths for the same question and returns the answer that the majority of those paths agree on

Prompt EngineeringReasoning Models

ZAYA1-8B

ZAYA1-8B is an open-weight, reasoning-focused Mixture-of-Experts (MoE) large language model released by San Francisco-based AI research lab Zyphra on May 6, 2026.

AI ModelsMixture of Experts

o4-mini

OpenAI o4-mini is a compact reasoning model developed by OpenAI, released on April 16, 2025.

AI ModelsOpenAI