DeepSeek V4.1-Flash
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
Explore language models, how they work, and the techniques used to build applications with them.
Articles that also belong to these categories. Counts cover all of Large Language Models.
Showing 121-180 of 545 articles
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
DeepSeek, Llama, and Qwen are families of large language models, not single systems. DeepSeek announced V4 as a preview on April 24, 2026 .
DeepSeek-Coder is a family of open-weight code large language models built for code generation, completion, and infilling, developed by the Chinese AI lab DeepSeek (DeepSeek-AI).
DeepSeek-Prover is a family of open-weight large language models developed by Chinese AI laboratory DeepSeek for formal theorem proving in the Lean 4 proof assistant.
DeepSeek-R1 is an open-weight reasoning model and large language model family developed by the Chinese artificial intelligence laboratory DeepSeek. The original model was released on January 20, 2025.
DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.
DeepSeek-V2 is a 236-billion-parameter mixture-of-experts (MoE) large language model released in May 2024 by DeepSeek, the Chinese AI lab spun out of the quantitative hedge fund High-Flyer and led by Liang…
DeepSeekMath is a family of open-weight large language models specialized for mathematical reasoning, released by Chinese AI laboratory DeepSeek in February 2024.
Devstral is a family of open-weight and API large language models specialized for agentic software engineering, developed by Mistral AI in collaboration with All Hands AI
Diffusion language models (DLMs, sometimes written dLLMs at frontier scale) are text generators that synthesize a sequence by iteratively denoising or unmasking many tokens in parallel
DiffusionGemma is an experimental open-weight large language model developed by Google DeepMind for multimodal text generation through discrete diffusion.
Dolma is an open three-trillion-token English pretraining corpus released by the Allen Institute for AI (AI2) to power its fully open OLMo language models and to let researchers study how training data shapes…
Doubao (豆包, literally "bean bun") is China's most-used artificial intelligence chatbot, developed by ByteDance, the parent company of TikTok.
Doubao Seed 1.6 is a family of general-purpose foundation models developed by the ByteDance Seed research team and released through Volcano Engine on 11 June 2025 at the company's Force Original Power…
EAGLE-2 ("Faster Inference of Language Models with Dynamic Draft Trees") is the second generation of the EAGLE family of speculative decoding methods for accelerating large language model inference, introduced…
ERNIE 4.5 is a family of large language models released by the Chinese technology company Baidu, open-sourced on June 30, 2025 under the Apache 2.0 license .
ERNIE 5.0 is a natively omni-modal foundation model from Baidu, unveiled at the company's annual Baidu World 2025 conference in Beijing on 13 November 2025 as the flagship in the ERNIE line at launch.
ERNIE X1 is a deep-reasoning large language model developed by Baidu, the Chinese search and artificial-intelligence company, as part of its ERNIE (Wenxin) family.
EXAONE (an acronym for EXpert AI for EveryONE) is the family of large language models and foundation models developed by LG AI Research
EleutherAI is a non-profit artificial intelligence research institute that builds and openly releases large language models, datasets, and evaluation tools, and studies their interpretability and alignment.
EmbeddingGemma is an open text embedding model from Google, released in September 2025, that turns text into dense numeric vectors for search, retrieval, classification, and clustering.
Emergent abilities are capabilities of large language models (LLMs) that are absent in smaller models but appear once a model reaches sufficient scale.
FACTS Grounding is a factuality benchmark from Google DeepMind and Google Research that measures whether a large language model answers a request using only the information in a provided source document
Falcon is a family of open-source large language models built by the Technology Innovation Institute (TII)
Falcon 3 is a family of open-weight large language models released on December 17, 2024 by the Technology Innovation Institute (TII), an applied research center based in Abu Dhabi, United Arab Emirates .
Falcon-H1 is a family of open-weight large language models released in 2025 by the Technology Innovation Institute (TII)
FineWeb-Edu is an open, English-language pretraining dataset of roughly 1.3 trillion tokens
Fireworks AI is an artificial intelligence infrastructure company that runs a high-performance inference platform for deploying and serving open large language models (LLMs), image generation models, audio…
A foundation model is a machine learning model trained on broad data, generally with self-supervision at scale, that can be adapted to a wide range of downstream tasks.
Frontier models are artificial intelligence models at or near a selected boundary of capability, scale, or risk.
Function calling is a capability of large language models (LLMs) that lets a model decide, during generation, to invoke an external function or API and emit the function name plus its arguments as structured…
GGUF (GPT-Generated Unified Format) is the standard binary file format for storing large language models for local inference, bundling a model's weights, tokenizer, and metadata into a single self-contained…
GLM, short for General Language Model, is both a pretraining framework for language understanding and generation and the name of a model family developed by researchers associated with Tsinghua University and…
GLM-130B is a 130-billion-parameter bilingual (English and Chinese) large language model released in August 2022 by the Knowledge Engineering Group (KEG) and the Data Mining research group at Tsinghua…
GLM-4 is the fourth-generation foundation model family from Zhipu AI, a Beijing company spun out of the Knowledge Engineering Group at Tsinghua University and now trading internationally as Z.ai.
GLM-4.5 is an open-weights large language model released by Zhipu AI (operating internationally as Z.ai) on July 28, 2025, built on a 355-billion-parameter Mixture of Experts architecture that activates 32…
GLM-4.6 is a flagship open-weight large language model released by Zhipu AI under its international brand Z.ai on September 30, 2025, built on a sparse Mixture of Experts (MoE) architecture with roughly 357…
GLM-5 is an open-weight flagship large language model released by the Chinese AI company Zhipu AI, under its international brand Z.ai, on February 11, 2026.
GLM-5.1 is an open-weight large language model developed by the Chinese AI company Zhipu AI, which markets its products internationally under the brand Z.ai.
GLM-5.2 is an open-weight large language model developed by the Chinese company Zhipu AI, which sells its products internationally under the brand Z.ai.
GLM-5.3 is a text-input and text-output large language model from Z.ai, available through hosted services and downloadable checkpoints.
GLM-5.3-Flash is an open-weight, natively multimodal mixture-of-experts large language model released by Z.ai on August 26, 2026.
GPT, short for Generative Pre-trained Transformer, is the name of a model family developed by OpenAI.
OpenAI's GPT lineage runs from GPT-1 (June 2018, 117 million parameters, a 512-token context window) to the GPT-5.6 "Sol, Terra, Luna" family that entered limited preview on June 26, 2026.
GPT-1 is the first model in the GPT (Generative Pre-trained Transformer) series, a 117-million-parameter, 12-layer decoder-only Transformer released by OpenAI on June 11, 2018 in the paper "Improving Language…
GPT-2 is a family of autoregressive language models introduced by OpenAI on February 14, 2019.
GPT-3 (Generative Pre-trained Transformer 3) is a family of decoder-only, autoregressive large language models developed by OpenAI.
GPT-3.5 is a family of large language models developed by OpenAI. OpenAI used the name for the series from which it fine-tuned the model behind the original ChatGPT research preview, launched on November 30
GPT-4 (Generative Pre-trained Transformer 4) is a large language model developed by OpenAI and released on March 14, 2023.
GPT-4 Turbo is a family of large language models released by OpenAI as a faster, cheaper, and longer-context variant of GPT-4.
GPT-4.1 is a family of multimodal large language models developed by OpenAI and announced on April 14, 2025, with a 1 million token context window and a SWE-bench Verified coding score of 54.6%
GPT-4.1 mini is a large language model developed by OpenAI and released on April 14, 2025 as the mid-size member of the GPT-4.1 family.
GPT-4.1 nano is the smallest, fastest, and least expensive model in OpenAI's GPT-4.1 family of large language models, released on April 14, 2025.
GPT-4.5 is a large language model developed by OpenAI and released as a research preview on February 27, 2025, then deprecated and removed from the API on July 14, 2025, less than five months later.
GPT-4V, also written GPT-4V(ision) and read as "GPT-4 with vision," is the image-understanding capability that OpenAI added to its GPT-4 large language model, letting a user supply one or more images alongside…
GPT-5 is a family of proprietary large language models released by OpenAI on August 7, 2025.
GPT-5-Codex is a coding-specialised variant of OpenAI's GPT-5 model, announced on September 15, 2025 and tuned for agentic software engineering inside the OpenAI Codex product family.
GPT-5 Pro is a large language model developed by OpenAI and the highest-capability variant of the GPT-5 model family.
GPT-5.1 is a family of large language models developed by OpenAI and released on November 12, 2025.
GPT-5.1-Codex-Max is a frontier agentic coding model from OpenAI, released on November 19, 2025 for Codex, OpenAI's software engineering agent.