AI Models

Explore AI Models through related topics and the articles other pages reference most.

Explore articles

Browse subtopics (64)

Articles that also belong to these categories. Counts cover all of AI Models.

Showing 301-360 of 408 articles

Phi-4

Phi-4 is a 14-billion-parameter small language model developed by Microsoft Research and released in December 2024, designed to match or beat models several times its size on reasoning tasks by training…

Large Language ModelsMicrosoft

Pika 2.5

Pika 2.5 is a generative video model developed by Pika Labs, the San Francisco based AI video startup co-founded in April 2023 by Stanford AI Lab dropouts Demi Guo and Chenlin Meng.

Computer VisionGenerative AI

Question Answering Models

Question answering (QA) models are natural language processing systems that take a natural-language question as input and return a natural-language answer, optionally grounded in a supplied passage, document…

Natural Language Processing

Qwen-Image

Qwen-Image is an open-weight image-generation foundation model released by Alibaba's Qwen team in August 2025.

Generative AI

Qwen-Image-3.0

Qwen-Image-3.0 is a text-to-image foundation model announced by Alibaba's Qwen team on July 21, 2026, as the third generation of the Qwen-Image series .

Chinese AIGenerative AI

Qwen3-Max

Qwen3-Max is the flagship large language model in Alibaba's Qwen series and the first Qwen model to cross one trillion parameters, released in preview on September 5, 2025 and formally launched at the Apsara…

Chinese AILarge Language Models

RT-2

RT-2 (Robotic Transformer 2) is a vision-language-action model developed by Google DeepMind that enables robots to execute novel tasks by transferring knowledge from internet-scale vision-language pretraining…

Google DeepMindRobotics

Recraft V3

Recraft V3 is a text-to-image generation model developed by Recraft AI and released on October 30, 2024, that became the first model to reach the number-one position on the Artificial Analysis Text-to-Image…

Generative AIImage Generation

Reka Core

Reka Core is a frontier class multimodal foundation model developed by Reka AI, a research and product company founded in 2022 by former scientists from DeepMind, Google Brain, Meta FAIR, and Baidu.

Large Language ModelsMultimodal AI

Reka Edge

Reka Edge is a 7-billion-parameter multimodal language model developed by Reka AI, introduced in April 2024 as the smallest member of the company's first publicly described model family.

Large Language ModelsMultimodal AI

Reka Flash

Reka Flash is a family of multimodal large language models developed by Reka AI, a San Francisco Bay Area research company founded in 2022 by former researchers from Google DeepMind, Meta FAIR, and Google.

Large Language ModelsMultimodal AI

Robot foundation model

A robot foundation model is a large-scale machine learning model, typically based on the transformer architecture, that is pre-trained on broad, diverse datasets of robot interactions and then adapted to a…

Robotics

Rodin Gen-2

Rodin Gen-2 is a generative artificial intelligence model for producing three dimensional assets from text prompts or reference images, developed by the Shanghai based company Deemos and delivered through the…

Chinese AIGenerative AI

Runway Aleph

Runway Aleph is an in-context AI video editing model from Runway that transforms and edits an existing video clip from a plain-text instruction, performing a wide range of tasks in a single model: adding…

Computer VisionGenerative AI

SAM 2

SAM 2 (Segment Anything Model 2) is a promptable visual segmentation model for both images and video developed by Meta AI and released on 29 July 2024.

Computer VisionMeta AI

Sapiens (computer vision)

Sapiens is a family of human-centric computer vision foundation models developed by Meta (Reality Labs), introduced in 2024 and presented as an oral paper at the European Conference on Computer Vision (ECCV)…

Computer VisionMeta AI

Seedance

Seedance is the family of foundation video generation models built by the Seed team at ByteDance, the Chinese internet company that owns TikTok and Douyin.

Chinese AIComputer Vision

Seedance 2.0

Seedance 2.0 is a multimodal video generation model developed by ByteDance, released in February 2026 as the second major version of the company's Seedance line.

Chinese AIVideo Generation

Seedance 2.5

Seedance 2.5 is a proprietary AI video generation model released by ByteDance on July 31, 2026. It jointly generates audio and video from combinations of text, image, video, and audio inputs.

Chinese AIComputer Vision

Seedream

Seedream is a series of text-to-image and image-editing foundation models built by the Seed research team at ByteDance, the company behind TikTok and Douyin.

Chinese AIComputer Vision

Seedream 5.0

Seedream 5.0 is a text-to-image generation model developed by ByteDance, released in February 2026 as the fifth major version of the company's Seedream line.

Chinese AIImage Generation

Sentence-BERT (SBERT)

Sentence-BERT (SBERT) is a modification of the pretrained BERT transformer network that produces semantically meaningful, fixed-size sentence embeddings comparable with simple cosine similarity.

Natural Language Processing

Sesame CSM

Sesame CSM (Conversational Speech Model) is an open weights speech generation model from Sesame AI, a San Francisco startup co-founded by former Oculus chief executive Brendan Iribe.

Generative AIOpen Source AI

Shap-E

Shap-E is a conditional generative model for 3D assets developed by OpenAI, introduced in the paper "Shap-E: Generating Conditional 3D Implicit Functions" by Heewoo Jun and Alex Nichol, submitted to arXiv on…

Generative AIOpenAI

Skywork R1V

Skywork R1V is a family of open-weight multimodal vision-language models built for chain-of-thought reasoning, developed by Skywork AI

Large Language Models

SmolLM

SmolLM is a family of small, fully open language models released by Hugging Face on July 16, 2024 in three sizes, 135 million, 360 million, and 1.7 billion parameters, all trained on a curated open dataset…

Large Language ModelsOpen Source AI

SmolLM 3

SmolLM 3 is a fully open 3 billion parameter language model released by Hugging Face on July 8, 2025, trained on 11.2 trillion tokens and designed as a small, multilingual, long-context reasoner.

Large Language ModelsOpen Source AI

SmolVLA

SmolVLA (Small Vision-Language-Action) is a compact, open-source vision-language-action model (VLA) for robotics developed by Hugging Face and released in June 2025.

AI HardwareArtificial Intelligence

Sora 2

Sora 2 is a text-to-video and audio generation model developed by OpenAI, released on September 30, 2025, that OpenAI called "the GPT-3.5 moment for video." It succeeded the original Sora research preview from…

Generative AIMultimodal AI

Step-3

Step-3 is an open-weight large multimodal mixture of experts (MoE) model released in July 2025 by StepFun, the Shanghai-based Chinese artificial intelligence startup also known as Jieyue Xingchen.

Large Language Models

Summarization Models

Summarization models are natural language processing systems that condense a source document, set of documents, or dialogue into a shorter version that preserves the most important information.

Natural Language Processing

Suno v5

Suno v5 is the fifth-generation AI music generation model from Suno Inc., the Cambridge, Massachusetts startup, released on September 23, 2025 to Pro and Premier subscribers as what Suno called "the world's…

Generative AIMusic & Audio Generation