Large Language Models

Explore language models, how they work, and the techniques used to build applications with them.

Explore articles

Browse subtopics (65)

Articles that also belong to these categories. Counts cover all of Large Language Models.

Showing 241-300 of 545 articles

Hunyuan

Hunyuan is Tencent's brand for its foundation models, an umbrella covering large language models, video generation, image synthesis, 3D asset creation, and multimodal reasoning.

AI ModelsChinese AI

IBM Granite

IBM Granite is a family of open foundation models built by IBM for enterprise use, released under the permissive Apache 2.0 license and delivered through IBM's watsonx platform.

AI ModelsOpen Source AI

IBM watsonx

IBM watsonx is an enterprise artificial intelligence and data platform built by IBM and announced on May 9, 2023, at IBM's Think conference by CEO Arvind Krishna.

AI CompaniesEnterprise AI

IFEval

IFEval (Instruction-Following Evaluation) is a benchmark of 541 prompts that measures how reliably large language models obey explicit, machine-checkable instructions such as "write in more than 400 words,"…

AI BenchmarksNatural Language Processing

Inception Labs

Inception Labs (often referred to simply as Inception) is a Palo Alto, California-based artificial intelligence startup that commercializes diffusion language models (dLLMs) for text and code generation.

AI CompaniesDiffusion Models

Indirect prompt injection

Indirect prompt injection is a class of attack against large language model-integrated applications in which the malicious instructions that subvert the model are not supplied by the user, but are smuggled…

AI Safety

Inference optimization

Inference optimization is the set of techniques that make running a trained artificial intelligence model, especially a large language model, faster, more memory-efficient, and cheaper to serve in production.

Machine Learning

InfiniteBench

InfiniteBench (stylized as ∞Bench) is a long-context benchmark that tests whether large language models (LLMs) can genuinely process and reason over inputs longer than 100,000 tokens, using 12 tasks that span…

AI BenchmarksNatural Language Processing

Inflection AI

Inflection AI is an American artificial intelligence company founded in March 2022 by Mustafa Suleyman (co-founder of Google DeepMind), Karén Simonyan, and Reid Hoffman (co-founder of LinkedIn).

AI CompaniesArtificial Intelligence

InstructGPT

InstructGPT is a family of language models released by OpenAI in January 2022 that take the base GPT-3 and fine-tune it to follow user instructions more helpfully, truthfully, and with less toxic output, using…

AI AlignmentOpenAI

InternLM

InternLM (Chinese name Shusheng Puyu, 书生·浦语) is a family of open-weight large language models developed primarily by Shanghai AI Laboratory, with partners that include SenseTime, the Chinese University of Hong…

Chinese AIOpen Source AI

InternVL3

InternVL3 is an open-weights family of multimodal AI large language models released on April 11, 2025 by OpenGVLab, the general vision team associated with Shanghai AI Laboratory.

AI Models

Jamba

Jamba is a family of open-weight large language models from AI21 Labs, first released on March 28, 2024, and is the world's first production-grade language model built on a Mamba state space model (SSM)…

AI CompaniesMixture of Experts

Jamba2

Jamba2 is the second generation of hybrid State Space Model and Transformer language models released by AI21 Labs on January 8, 2026.

AI ModelsMixture of Experts

KV cache offloading

KV cache offloading is the practice of moving part or all of a Transformer model's KV cache out of accelerator memory (GPU HBM) into a larger, slower tier such as CPU DRAM, local NVMe storage, or remote…

AI InferenceAI Infrastructure

Kimi K1.5

Kimi K1.5 is a multimodal reasoning large language model developed by Moonshot AI, a Beijing-based artificial intelligence company.

AI ModelsChinese AI

Kimi K2

Kimi K2 is an open-weights Mixture of Experts language model from Moonshot AI, a Beijing startup, released on July 11, 2025 with 1.04 trillion total parameters and 32.6 billion activated per token

AI ModelsChinese AI

Kimi K2.5

Kimi K2.5 is an open-weights, natively multimodal large language model developed by Moonshot AI and released on January 27, 2026 .

AI ModelsChinese AI

Kimi K2.6

Kimi K2.6 is an open-weight, trillion-parameter mixture of experts (MoE) large language model released by Moonshot AI on 20 April 2026 for agentic coding and long-horizon autonomous execution.

Chinese AIOpen Source AI

Kimi K3

Kimi K3 is an open-weight multimodal reasoning model developed by Moonshot AI. Moonshot made the model available through its hosted products on July 16, 2026 and published the weights, code, configuration…

AI ModelsChinese AI

Kimi Linear

Kimi Linear is a hybrid linear attention architecture published by Moonshot AI on October 30, 2025, together with a 48-billion-parameter mixture-of-experts model that activates 3 billion parameters per token.

AI ModelsChinese AI

LLM Compiler (Meta)

The Meta Large Language Model Compiler, usually shortened to LLM Compiler, is a family of pre-trained large language model models built by Meta AI for code and compiler optimization tasks.

AI Code GenerationMeta AI

LLM Evaluation

LLM evaluation is the practice of measuring what a large language model can do, how reliably it does it, and how it behaves under adversarial or high-stakes conditions.

AI BenchmarksAI Research

LLM inference engine

An LLM inference engine (also called an LLM serving engine or LLM inference server) is the systems software stack that loads trained large language model weights into GPU or CPU memory and answers user…

AI InferenceAI Infrastructure

LLM.int8()

LLM.int8() is an 8-bit matrix multiplication scheme for large language model inference that preserves accuracy across models up to 175 billion parameters by combining vector-wise quantization with a…

AI Inference

LLaDA (Large Language Diffusion)

LLaDA (Large Language Diffusion with mAsking) is a family of non-autoregressive large language models that generate text by iteratively denoising a sequence of mask tokens rather than predicting tokens left to…

Diffusion Models

LM Studio

LM Studio is a desktop application for discovering, downloading, and running large language models locally on personal hardware, available free for both personal and commercial use on macOS, Windows, and Linux.

Developer ToolsOpen Source AI

LaMDA

LaMDA (short for Language Model for Dialogue Applications) is a family of conversational large language models developed by Google, built on the Transformer architecture and fine-tuned for open-ended dialogue…

Conversational AIGoogle

LayoutLM

LayoutLM is a family of pre-trained multimodal models developed by Microsoft Research for document AI, the task of automatically reading and understanding visually rich documents such as forms, invoices…

Multimodal AI

Ling-1T

Ling-1T is a trillion-parameter open-weight language model released by Ant Group through its Inclusion AI research group on October 9, 2025.

Chinese AIMixture of Experts

Ling-3.0-flash

Ling-3.0-flash is an open-weight mixture-of-experts language model from inclusionAI, the open-source AI initiative of Ant Group, and the first model of the Ling 3.0 generation.

AI ModelsChinese AI