AI Models

Explore AI Models through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: Robotics

Articles that also belong to these categories. Counts cover all of AI Models.

Showing 1-29 of 29 articles

Gemini Robotics

Gemini Robotics is a family of robot foundation models developed by Google DeepMind that extends the Gemini multimodal model line into the physical world.

Google DeepMindRobotics

Gemini Robotics 2

Gemini Robotics 2 is a family of three robotics models announced by Google DeepMind on July 30, 2026: a vision-language-action (VLA) model of the same name, an embodied reasoning model called Gemini Robotics…

Embodied AIGoogle DeepMind

Helix (Figure AI)

Helix is a Vision-Language-Action (VLA) model for generalist humanoid control, developed in-house by Figure AI, the Sunnyvale, California humanoid-robot company founded by Brett Adcock.

Embodied AIRobotics

Helix (VLA model)

Helix is a vision-language-action model (VLA) developed by Figure AI that controls humanoid robots by mapping camera images and natural language commands directly to continuous joint-level motion at 200 Hz.

Embodied AIHumanoid Robots

Hydra-0

Hydra-0 is an experimental world model for robot manipulation that represents actions as trajectories of visible points in an image.

NVIDIARobotics

Lumo-2

Lumo-2 is a 4B-scale robot-learning model developed by Astribot. Astribot describes it as a latent world-action model: instead of rendering a future video before acting, it predicts an action-relevant…

Embodied AIPhysical AI

NVIDIA COMPASS

NVIDIA COMPASS is a framework and trained model for cross-embodiment robot navigation. Its name expands to Cross-Embodiment Mobility Policy via Residual RL and Skill Synthesis.

Embodied AINVIDIA

NVIDIA Isaac GR00T N1

NVIDIA Isaac GR00T N1 is an open foundation model for humanoid robots developed by NVIDIA and unveiled by Jensen Huang on March 18, 2025 at the company's annual GTC conference in San Jose, California.

Humanoid RobotsOpen Source AI

OpenPI

OpenPI (stylized openpi) is the open-source repository of robot foundation models, training code, and inference utilities published by Physical Intelligence, the San Francisco robotics and AI startup…

Developer ToolsOpen Source AI

OpenVLA

OpenVLA is a 7-billion-parameter open-source vision-language-action model (VLA) for robotic manipulation, released in June 2024 by a collaboration of researchers from Stanford University, UC Berkeley, the…

Open Source AIRobotics

RT-2

RT-2 (Robotic Transformer 2) is a vision-language-action model developed by Google DeepMind that enables robots to execute novel tasks by transferring knowledge from internet-scale vision-language pretraining…

Google DeepMindRobotics

Robot foundation model

A robot foundation model is a large-scale machine learning model, typically based on the transformer architecture, that is pre-trained on broad, diverse datasets of robot interactions and then adapted to a…

Robotics

SmolVLA

SmolVLA (Small Vision-Language-Action) is a compact, open-source vision-language-action model (VLA) for robotics developed by Hugging Face and released in June 2025.

AI HardwareArtificial Intelligence

V-JEPA 2

V-JEPA 2 (Video Joint Embedding Predictive Architecture 2) is an open-source video world model released by Meta AI on June 11, 2025 that learns to understand, predict, and plan in the physical world by…

Computer VisionOpen Source AI

World action model

A world action model (WAM) is a robot policy design that builds action generation on a video world model backbone rather than on a vision-language model, so that a single network jointly predicts how a scene…

Embodied AIRobotics

π0

π0 (pronounced "pi-zero") is a vision-language-action model for general-purpose robot control developed by Physical Intelligence, a San Francisco-based robotics startup, and introduced on October 31, 2024.

Robotics

π0.5

π0.5 (also written pi0.5, pi 0.5, or π₀.₅, and pronounced "pi zero point five") is a vision-language-action model developed by the robotics company Physical Intelligence and released on April 22, 2025.

Embodied AIRobotics

π₀ (pi-zero)

π₀ (pronounced pi-zero and sometimes written pi0 or pizero) is a vision-language-action model (VLA) developed by the robotics foundation-model startup Physical Intelligence

Embodied AIRobotics