AI Models

Explore AI Models through related topics and the articles other pages reference most.

Explore articles

Browse subtopics (64)

Articles that also belong to these categories. Counts cover all of AI Models.

Showing 61-120 of 408 articles

DALL-E 2

DALL-E 2 (stylized DALL·E 2) is a text-to-image generation system that OpenAI announced on April 6, 2022, capable of producing photorealistic 1024 by 1024 pixel images from a written prompt

Image GenerationOpenAI

DALL-E 3

DALL-E 3 (stylized by OpenAI as DALL·E 3) is the third generation of OpenAI's text-to-image system, announced on September 20, 2023 and released to ChatGPT Plus and Enterprise subscribers in October 2023

Image GenerationOpenAI

DINOv2

DINOv2 is a family of self-supervised Vision Transformer models released by Meta AI Research in April 2023 that produces general-purpose visual features transferring to many downstream tasks without…

Computer VisionMeta AI

DeepSeek Janus

DeepSeek Janus is a family of open-weight unified multimodal models from Chinese AI lab DeepSeek that perform both image understanding and text-to-image generation in a single autoregressive Transformer.

Chinese AIMultimodal AI

DeepSeek V3

DeepSeek-V3 is a 671-billion-parameter open-weights Mixture of Experts large language model from Chinese AI lab DeepSeek, released on December 26, 2024, that activates only 37 billion parameters per token and…

Chinese AILarge Language Models

DeepSeek V4-Flash

DeepSeek V4-Flash is the smaller of the two large language models in the DeepSeek V4 family, a 284-billion-parameter Mixture of Experts model with 13 billion active parameters and a one-million-token context…

AI Code GenerationChinese AI

DeepSeek V4-Pro

DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…

Chinese AILarge Language Models

DistilBERT

DistilBERT is a compressed version of BERT released by Hugging Face in October 2019 that is 40% smaller and 60% faster than BERT-base while retaining 97% of its language-understanding performance on the GLUE…

Deep LearningNatural Language Processing

DistilGPT2

DistilGPT2 is an English-language autoregressive language model released by Hugging Face on October 3, 2019.

Document Question Answering Models

Document question answering models (DocQA, sometimes called DocVQA for document visual question answering) are machine learning systems that take a document image or PDF together with a natural language…

Multimodal AI

Donut (Model)

Donut (Document understanding transformer) is an OCR-free visual document understanding model introduced by researchers at NAVER CLOVA in the paper "OCR-free Document Understanding Transformer," first posted…

Computer VisionMultimodal AI

Doubao Seed 1.6

Doubao Seed 1.6 is a family of general-purpose foundation models developed by the ByteDance Seed research team and released through Volcano Engine on 11 June 2025 at the company's Force Original Power…

Chinese AILarge Language Models

Doubao Seedance

Doubao-Seedance is a family of video generation foundation models developed by ByteDance Seed, the AI research division of Chinese technology conglomerate ByteDance.

Chinese AIVideo Generation

Doubao Seedream

Doubao-Seedream is the family of text-to-image generation foundation models developed by the ByteDance Seed team and shipped through ByteDance's Doubao product line and the company's Volcano Engine cloud…

Chinese AIImage Generation

ESM3

ESM3 (Evolutionary Scale Modeling 3) is a frontier multimodal generative language model for biology, released by EvolutionaryScale on June 25, 2024, that was the first model to reason jointly over the…

Drug DiscoveryHealthcare AI

ESMFold

ESMFold is a protein structure prediction model developed by the Meta AI Fundamental AI Research (FAIR) Protein Team.

AI for ScienceMeta AI

ElevenLabs v3

Eleven v3, marketed by ElevenLabs as Eleven v3 (alpha), is a third-generation text-to-speech model that ElevenLabs released in public alpha on June 5, 2025 and described as "the most expressive Text to Speech…

Generative AISpeech & Audio AI

Emu Edit

Emu Edit is an instruction-based image editing model from Meta AI, announced on November 16, 2023 alongside the text-to-video model Emu Video.

Image GenerationMeta AI

Emu Video

Emu Video is a text-to-video generation model from Meta AI, announced on November 16, 2023, that creates short clips by first turning a text prompt into an image and then generating a video conditioned on both…

Meta AIVideo Generation

Extropic Z1T

Z1T is a family of sparse, transformer-like language models that Extropic designed to run partly on its Z1 probabilistic chip, described in a research post dated September 4, 2026 by Guillaume Verdon…

AI HardwareOpen Source AI

F5-TTS

F5-TTS (short for "A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching") is an open-source text-to-speech and zero-shot voice cloning model released in October 2024 by researchers from…

Open Source AISpeech & Audio AI

Falcon 3

Falcon 3 is a family of open-weight large language models released on December 17, 2024 by the Technology Innovation Institute (TII), an applied research center based in Abu Dhabi, United Arab Emirates .

AI CompaniesLarge Language Models

Feature Extraction Models

Feature extraction models are machine learning systems that transform raw inputs such as text, images, or audio into dense numerical vectors known as embeddings or hidden-state representations.

Multimodal AI

Fill-Mask Models

Fill-mask models are language models trained with a masked language modeling (MLM) objective, in which a fraction of the tokens in an input sequence are hidden behind a special [MASK] symbol and the model…

Natural Language Processing

Florence-2

Florence-2 is a vision foundation model developed by Microsoft Research that handles a wide range of computer vision and vision-language tasks through a single unified

Computer VisionMicrosoft

Frontis-MA1

Frontis-MA1 is a family of open-weight large language models post-trained to act as agents for machine learning engineering (MLE), released in late July 2026 by FrontisAI

AI AgentsChinese AI

GAIA-4 (Wayve)

GAIA-4 is a multimodal generative world model for closed-loop autonomous driving simulation, announced by the British self-driving company Wayve on 3 August 2026 as the latest generation of its GAIA family.

Autonomous VehiclesComputer Vision