AI Models

Explore AI Models through related topics and the articles other pages reference most.

Explore articles

Browse subtopics (64)

Articles that also belong to these categories. Counts cover all of AI Models.

Showing 361-408 of 408 articles

Synthesia 3.0

Synthesia 3.0 is a major release of the AI video generation platform from Synthesia, the London-based company co-founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Lourdes Agapito and Matthias Niessner.

Computer VisionGenerative AI

Tabular Regression Models

Tabular regression models are machine learning systems that predict a continuous numeric target from a vector of tabular features, where rows are samples and columns are heterogeneous attributes (numeric…

Machine Learning

Tripo P1

Tripo P1, marketed in full as Tripo Smart Mesh P1.0, is a production grade native 3D diffusion model that generates clean, engine ready 3D meshes from text or image prompts in as little as two seconds.

Chinese AIGenerative AI

Unconditional Image Generation Models

Unconditional image generation models are generative neural networks that learn the marginal distribution p(x) of a set of training images and produce new samples from that learned distribution, with no extra…

Computer Vision

V-JEPA

V-JEPA (Video Joint Embedding Predictive Architecture) is a self-supervised video model from Meta AI that learns by predicting masked regions of a video in an abstract latent representation space rather than…

Computer VisionMachine Learning

V-JEPA 2

V-JEPA 2 (Video Joint Embedding Predictive Architecture 2) is an open-source video world model released by Meta AI on June 11, 2025 that learns to understand, predict, and plan in the physical world by…

Computer VisionOpen Source AI

VGGNet

VGGNet is a deep convolutional neural network architecture, introduced in 2014 by Karen Simonyan and Andrew Zisserman of the Visual Geometry Group at the University of Oxford, that classifies images using a…

AI HistoryComputer Vision

Veo 3

Veo 3 is a video generation model developed by Google DeepMind and announced at Google I/O on May 20, 2025, and it is the first commercially available video generation model to natively produce synchronized…

Google DeepMindVideo Generation

Voice Activity Detection Models

Voice activity detection (VAD), also called speech activity detection (SAD), is the task of deciding which segments of an audio signal contain human speech and which contain only silence, background noise…

Speech & Audio AI

Voyage-3

Voyage-3 is a family of general-purpose text embedding models developed by Voyage AI, launched in September 2024 with voyage-3 and voyage-3-lite , expanded in January 2025 with voyage-3-large , and refreshed…

AnthropicInformation Retrieval

Wan 2.1

Wan 2.1 (also written Wan2.1, from the Chinese Tongyi Wanxiang or 通义万象) is a family of open-weights text-to-video and image-to-video generation models that Alibaba's Tongyi Wanxiang team released and…

Chinese AIGenerative AI

Wan 2.5

Wan 2.5 is a natively multimodal AI video generation model developed by Alibaba Cloud's Tongyi Lab and previewed at the company's Apsara 2025 conference in Hangzhou on September 24, 2025.

Chinese AIComputer Vision

WeatherNext 2

WeatherNext 2 is an artificial intelligence weather forecasting model from Google DeepMind and Google Research, announced on November 17, 2025

AI for Science

World action model

A world action model (WAM) is a robot policy design that builds action generation on a video world model backbone rather than on a vision-language model, so that a single network jointly predicts how a scene…

Embodied AIRobotics

Yi-Large

Yi-Large is a closed-source large language model developed by Chinese artificial intelligence company 01.AI (零一万物, Língyi Wànwù), founded by Kai-Fu Lee.

Chinese AILarge Language Models

Yi-Lightning

Yi-Lightning is a closed-source large language model developed by Chinese artificial intelligence company 01.AI (零一万物, Língyī Wànwù), the company founded by Kai-Fu Lee.

Chinese AILarge Language Models

rBio (CZI)

rBio (styled rbio1 in the accompanying preprint) is a biological reasoning large language model released by the Chan Zuckerberg Initiative (CZI) on August 21, 2025.

AI for Science

π0

π0 (pronounced "pi-zero") is a vision-language-action model for general-purpose robot control developed by Physical Intelligence, a San Francisco-based robotics startup, and introduced on October 31, 2024.

Robotics

π0.5

π0.5 (also written pi0.5, pi 0.5, or π₀.₅, and pronounced "pi zero point five") is a vision-language-action model developed by the robotics company Physical Intelligence and released on April 22, 2025.

Embodied AIRobotics

π₀ (pi-zero)

π₀ (pronounced pi-zero and sometimes written pi0 or pizero) is a vision-language-action model (VLA) developed by the robotics foundation-model startup Physical Intelligence

Embodied AIRobotics