Large Language Models

Explore language models, how they work, and the techniques used to build applications with them.

Explore articles

Browse subtopics (65)

Articles that also belong to these categories. Counts cover all of Large Language Models.

Showing 361-420 of 545 articles

Moonshot AI

Moonshot AI is a Beijing-based artificial intelligence company that develops the Kimi chatbot, large language models, and agent software. The company says it was founded in early 2023.

AI CompaniesChinese AI

Multi-agent system

A multi-agent system (MAS) is a system composed of multiple interacting intelligent agents that collaborate, compete, or negotiate to accomplish tasks that would be difficult or impossible for a single agent.

Artificial Intelligence

Multi-hop RAG

Multi-hop RAG is a family of retrieval-augmented generation techniques designed to answer questions that require composing evidence from two or more documents or text chunks.

Information Retrieval

Muse Glimmer

Muse Glimmer is an open-weight text-and-image model developed by Meta AI for local agent and coding workloads. Meta released the model on August 10, 2026 under the identifier meta-models/Muse-Glimmer-30B.

AI AgentsAI Models

NVLM

NVLM (short for NVIDIA Vision Language Model), released as NVLM 1.0, is a family of open multimodal large language models developed by Nvidia.

Multimodal AINVIDIA

Naver AI

Naver AI refers to the artificial intelligence research and products developed by Naver Corporation, South Korea's largest internet company.

AI Companies

Needle in a Haystack (NIAH)

Needle in a Haystack (NIAH) is a long-context evaluation that measures whether a large language model can retrieve a single fact (the "needle") inserted at a controlled position inside a long body of text (the…

AI Benchmarks

Nemotron

Nemotron is NVIDIA's brand for its family of open large language models and the datasets, training recipes, and evaluation tools built around them.

AI ModelsNVIDIA

Nemotron 3

Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…

AI ModelsNVIDIA

Nemotron-4

Nemotron-4 is a family of decoder-only large language models developed by NVIDIA and documented in two technical reports released in 2024.

AI ModelsNVIDIA

Nemotron-H

Nemotron-H is a family of open-weight large language models released by NVIDIA in April 2025 that replace most of the self-attention layers of a standard Transformer with Mamba-2 state-space layers, producing…

AI ModelsNVIDIA

NoLiMa

NoLiMa, short for "No Literal Matching," is a long-context benchmark for large language models that measures how well a model can find and use a single relevant fact buried in a long document when that fact…

AI BenchmarksModel Evaluation

Nous Research

Nous Research is a New York City-based applied AI research organization and company that builds widely used open-weight language models and decentralized training infrastructure.

AI ResearchOpen Source AI

OLMo

OLMo (Open Language Model) is a family of fully open large language models built by the Allen Institute for AI (Ai2) and first released on February 1, 2024.

AI ModelsOpen Source AI

OLMo 2

OLMo 2 is the second generation of fully open large language models released by the Allen Institute for AI (Ai2), spanning 7B, 13B, and 32B parameter sizes.

AI ModelsOpen Source AI

OLMoE

OLMoE (Open Mixture-of-Experts) is a fully open sparse mixture of experts large language model released by the Allen Institute for AI (Ai2) on September 3, 2024 .

AI ModelsMixture of Experts

OPUS-MT

OPUS-MT is a large collection of open, freely licensed neural machine translation models and tools produced by the Language Technology Research Group at the University of Helsinki.

Natural Language Processing

ORPO

ORPO (Odds Ratio Preference Optimization) is a preference alignment algorithm for large language models that merges supervised fine-tuning and preference alignment into a single training stage, eliminating the…

Machine LearningTraining & Optimization

Ollama

Ollama is a free, open-source runtime for downloading, running, and managing open-weight large language models (LLMs) locally on personal computers and servers.

Developer ToolsOpen Source AI

OpenAI API

The OpenAI API is a REST-based application programming interface that gives developers programmatic access to OpenAI's family of artificial intelligence models, including the GPT series of large language…

AI Tools & ProductsOpenAI

OpenAI o-series

The OpenAI o-series is a family of large language models developed by OpenAI that are trained with reinforcement learning to reason through an internal chain-of-thought before answering, making them OpenAI's…

Artificial IntelligenceOpenAI

OpenAI o1

OpenAI o1 is a family of proprietary large language models developed by OpenAI and trained to use additional computation before returning an answer.

AI ModelsOpenAI

OpenAI o1-mini

OpenAI o1-mini is a smaller, faster, and cheaper reasoning model released by OpenAI on September 12, 2024, alongside o1-preview, and optimized for science, technology, engineering, and mathematics (STEM) tasks…

OpenAIReasoning Models

OpenAI o1-pro

OpenAI o1-pro is the highest-compute variant of OpenAI's o1 reasoning model, designed to spend more inference-time compute so it "thinks harder" and returns the most reliable answers on the hardest…

OpenAIReasoning Models

OpenAI o3

OpenAI o3 is a family of reasoning-focused large language models developed by OpenAI and the second generation of the company's o-series reasoning models, best known for scoring 87.5% on the ARC-AGI…

AI ModelsOpenAI

OpenAI o3-mini

OpenAI o3-mini is a reasoning-focused large language model released by OpenAI on January 31, 2025, the second commercial member of the o-series after OpenAI o1 and a smaller, cheaper

OpenAIReasoning Models

OpenAI o3-pro

OpenAI o3-pro is a high-compute reasoning large language model released by OpenAI on June 10, 2025, designed as the professional, higher-reliability variant of the company's o3 reasoning model.

OpenAIReasoning Models

OpenOrca

OpenOrca is a large open-source instruction-tuning dataset that augments the FLAN Collection with chain-of-thought responses generated by OpenAI's GPT-3.5 and GPT-4 APIs.

Data & DatasetsOpen Source AI

OpenRouter

OpenRouter is a unified API gateway and marketplace that routes a single, OpenAI-compatible request across more than 400 large language models (LLMs) and other AI models from over 60 providers, automatically…

Artificial IntelligenceDeveloper Tools

Outlines (library)

Outlines is an open-source Python (programming language) library, released under the Apache 2.0 license, that constrains large language model output to user-specified structures: regular expressions, function…

Developer ToolsOpen Source AI

PaLM 2

PaLM 2 is the large language model Google announced on May 10, 2023, at its I/O developer conference as the successor to the original PaLM (Pathways Language Model).

Google

Patchscopes

Patchscopes is an interpretability framework for inspecting hidden representations of large language models by patching an internal activation from a source computation into a separate target inference whose…

Interpretability

Persona vectors

Persona vectors are single linear directions in the activation space of a large language model that correspond to high level character traits such as evil, sycophancy, or a propensity to hallucinate.

AI SafetyInterpretability

Phi (language model)

Phi is a family of open-weight small language models (SLMs) developed by Microsoft Research, beginning with Phi-1 in June 2023 and spanning thirteen-plus releases through Phi-4-reasoning-vision-15B in March…

MicrosoftOpen Source AI

Phi-4

Phi-4 is a 14-billion-parameter small language model developed by Microsoft Research and released in December 2024, designed to match or beat models several times its size on reasoning tasks by training…

AI ModelsMicrosoft

Phi-4-mini

Phi-4-mini is a 3.8 billion parameter open weight small language model released by Microsoft on February 26, 2025, under the permissive MIT license.

AI ModelsOpen Source AI

PiSSA

PiSSA (Principal Singular values and Singular vectors Adaptation) is a parameter-efficient fine-tuning method for large language models that initializes LoRA-style low-rank adapter matrices from the dominant…

Training & Optimization

Pixtral

Pixtral is a family of multimodal vision-language models developed by Mistral AI, a French AI company founded in April 2023.

AI CompaniesAI Models

Prompt

A prompt is the input given to a generative AI model, particularly a large language model (LLM), that elicits a desired response.

Prompt Engineering