Claude Sonnet 5
Claude Sonnet 5 is a large language model developed by Anthropic, released on June 30, 2026 as the mid-tier, default member of the current Claude lineup.
Explore AI Models through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Models.
Showing 61-120 of 408 articles
Claude Sonnet 5 is a large language model developed by Anthropic, released on June 30, 2026 as the mid-tier, default member of the current Claude lineup.
As of July 2026, neither Claude nor ChatGPT is universally better: the honest answer depends on the task.
Codestral is a family of code-specialized large language models developed by Mistral AI, beginning with Codestral 22B, released on May 29, 2024
Cohere Command A is a 111 billion parameter dense large language model released by Cohere on March 13, 2025, built for enterprise agents, Retrieval-Augmented Generation, and tool use across 23 languages.
Cohere Transcribe is a family of automatic speech recognition models developed by Cohere and Cohere Labs.
Command R+ is a large language model developed by Cohere, the Toronto- and San Francisco-based enterprise artificial intelligence company, and released on April 4, 2024.
Conversational models are computational systems designed to carry on a dialogue with human users in natural language
CrowdStrike SafeMind is a family of cybersecurity models and agent harnesses announced by CrowdStrike on September 1, 2026.
Cursor Composer 2.5 is a proprietary agentic coding model built by Anysphere, the company behind the Cursor code editor.
DALL-E 2 (stylized DALL·E 2) is a text-to-image generation system that OpenAI announced on April 6, 2022, capable of producing photorealistic 1024 by 1024 pixel images from a written prompt
DALL-E 3 (stylized by OpenAI as DALL·E 3) is the third generation of OpenAI's text-to-image system, announced on September 20, 2023 and released to ChatGPT Plus and Enterprise subscribers in October 2023
DINOv2 is a family of self-supervised Vision Transformer models released by Meta AI Research in April 2023 that produces general-purpose visual features transferring to many downstream tasks without…
DINOv3 is a family of self-supervised computer vision foundation models released by Meta AI in August 2025.
DeepSeek Janus is a family of open-weight unified multimodal models from Chinese AI lab DeepSeek that perform both image understanding and text-to-image generation in a single autoregressive Transformer.
DeepSeek LLM is the first foundational large language model series released by the Chinese AI company DeepSeek.
DeepSeek-V3 is a 671-billion-parameter open-weights Mixture of Experts large language model from Chinese AI lab DeepSeek, released on December 26, 2024, that activates only 37 billion parameters per token and…
DeepSeek V3.1 is a large language model developed by DeepSeek, released on August 19, 2025 and made broadly available via the official API on August 21, 2025.
DeepSeek V4 is a family of open-weight Mixture of Experts large language models developed by DeepSeek, a Hangzhou-based AI research lab.
DeepSeek V4-Flash is the smaller of the two large language models in the DeepSeek V4 family, a 284-billion-parameter Mixture of Experts model with 13 billion active parameters and a one-million-token context…
DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
DeepSeek, Llama, and Qwen are families of large language models, not single systems. DeepSeek announced V4 as a preview on April 24, 2026 .
DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.
DeepSeek-VL is the first open-source vision-language model series from DeepSeek, the Chinese AI company.
Deepgram Nova-3 is the third-generation automatic speech recognition (ASR) model developed by Deepgram, a San Francisco-based voice AI company.
Devstral is a family of open-weight and API large language models specialized for agentic software engineering, developed by Mistral AI in collaboration with All Hands AI
DistilBERT is a compressed version of BERT released by Hugging Face in October 2019 that is 40% smaller and 60% faster than BERT-base while retaining 97% of its language-understanding performance on the GLUE…
DistilGPT2 is an English-language autoregressive language model released by Hugging Face on October 3, 2019.
Document question answering models (DocQA, sometimes called DocVQA for document visual question answering) are machine learning systems that take a document image or PDF together with a natural language…
Donut (Document understanding transformer) is an OCR-free visual document understanding model introduced by researchers at NAVER CLOVA in the paper "OCR-free Document Understanding Transformer," first posted…
Doubao Seed 1.6 is a family of general-purpose foundation models developed by the ByteDance Seed research team and released through Volcano Engine on 11 June 2025 at the company's Force Original Power…
Doubao-Seedance is a family of video generation foundation models developed by ByteDance Seed, the AI research division of Chinese technology conglomerate ByteDance.
Doubao-Seedream is the family of text-to-image generation foundation models developed by the ByteDance Seed team and shipped through ByteDance's Doubao product line and the company's Volcano Engine cloud…
ELMo (Embeddings from Language Models) is a deep contextualized word embedding method, introduced in 2018 by the Allen Institute for AI (AI2) and the University of Washington
ERNIE X1 is a deep-reasoning large language model developed by Baidu, the Chinese search and artificial-intelligence company, as part of its ERNIE (Wenxin) family.
ESM3 (Evolutionary Scale Modeling 3) is a frontier multimodal generative language model for biology, released by EvolutionaryScale on June 25, 2024, that was the first model to reason jointly over the…
ESMFold is a protein structure prediction model developed by the Meta AI Fundamental AI Research (FAIR) Protein Team.
ElevenLabs Music (also marketed as Eleven Music) is an AI music generation product developed by ElevenLabs, the voice AI company founded in 2022 by Piotr Dabkowski and Mati Staniszewski.
Eleven v3, marketed by ElevenLabs as Eleven v3 (alpha), is a third-generation text-to-speech model that ElevenLabs released in public alpha on June 5, 2025 and described as "the most expressive Text to Speech…
Emu is a text-to-image generation foundation model developed by Meta AI and unveiled at the Meta Connect conference in September 2023.
Emu Edit is an instruction-based image editing model from Meta AI, announced on November 16, 2023 alongside the text-to-video model Emu Video.
Emu Video is a text-to-video generation model from Meta AI, announced on November 16, 2023, that creates short clips by first turning a text prompt into an image and then generating a video conditioned on both…
Evo 2 is a genomic foundation model built by the Arc Institute together with NVIDIA, Stanford University, and collaborators, and first released in February 2025.
Z1T is a family of sparse, transformer-like language models that Extropic designed to run partly on its Z1 probabilistic chip, described in a research post dated September 4, 2026 by Guillaume Verdon…
F5-TTS (short for "A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching") is an open-source text-to-speech and zero-shot voice cloning model released in October 2024 by researchers from…
FLUX 3 Video is a video generation model by Black Forest Labs (BFL), released into general availability on August 4, 2026.
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
FLUX.2 is the second-generation image generation and editing model family developed by Black Forest Labs, released on November 25, 2025.
Falcon 3 is a family of open-weight large language models released on December 17, 2024 by the Technology Innovation Institute (TII), an applied research center based in Abu Dhabi, United Arab Emirates .
Falcon-H1 is a family of open-weight large language models released in 2025 by the Technology Innovation Institute (TII)
Feature extraction models are machine learning systems that transform raw inputs such as text, images, or audio into dense numerical vectors known as embeddings or hidden-state representations.
Fill-mask models are language models trained with a masked language modeling (MLM) objective, in which a fraction of the tokens in an input sequence are hidden behind a special [MASK] symbol and the model…
Flamingo is a family of visual language models (VLMs) built by DeepMind and introduced in April 2022 that brought few-shot, in-context learning to multimodal inputs.
Florence-2 is a vision foundation model developed by Microsoft Research that handles a wide range of computer vision and vision-language tasks through a single unified
A foundation model is a machine learning model trained on broad data, generally with self-supervision at scale, that can be adapted to a wide range of downstream tasks.
Frontis-MA1 is a family of open-weight large language models post-trained to act as agents for machine learning engineering (MLE), released in late July 2026 by FrontisAI
GAIA-2 (Generative AI for Autonomy 2) is a controllable, multi-camera generative world model for autonomous driving, announced by the British self-driving company Wayve on 26 March 2025.
GAIA-3 is a 15-billion-parameter generative world model for autonomous driving released by Wayve on 2 December 2025, the third generation in the company's GAIA family.
GAIA-4 is a multimodal generative world model for closed-loop autonomous driving simulation, announced by the British self-driving company Wayve on 3 August 2026 as the latest generation of its GAIA family.
GEM (Generative Ads Recommendation Model) is a proprietary foundation model for advertising recommendation developed by Meta Platforms.