AI Models

Explore AI Models through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: Embodied AI

Articles that also belong to these categories. Counts cover all of AI Models.

Showing 1-17 of 17 articles

Gemini Robotics 2

Gemini Robotics 2 is a family of three robotics models announced by Google DeepMind on July 30, 2026: a vision-language-action (VLA) model of the same name, an embodied reasoning model called Gemini Robotics…

Embodied AIGoogle DeepMind

Helix (Figure AI)

Helix is a Vision-Language-Action (VLA) model for generalist humanoid control, developed in-house by Figure AI, the Sunnyvale, California humanoid-robot company founded by Brett Adcock.

Embodied AIRobotics

Helix (VLA model)

Helix is a vision-language-action model (VLA) developed by Figure AI that controls humanoid robots by mapping camera images and natural language commands directly to continuous joint-level motion at 200 Hz.

Embodied AIHumanoid Robots

Lumo-2

Lumo-2 is a 4B-scale robot-learning model developed by Astribot. Astribot describes it as a latent world-action model: instead of rendering a future video before acting, it predicts an action-relevant…

Embodied AIPhysical AI

Meta Motivo

Meta Motivo is a behavioral foundation model for controlling a simulated humanoid body, released by Meta AI's Fundamental AI Research (FAIR) group on December 12, 2024.

Embodied AIMeta AI

NVIDIA COMPASS

NVIDIA COMPASS is a framework and trained model for cross-embodiment robot navigation. Its name expands to Cross-Embodiment Mobility Policy via Residual RL and Skill Synthesis.

Embodied AINVIDIA

SmolVLA

SmolVLA (Small Vision-Language-Action) is a compact, open-source vision-language-action model (VLA) for robotics developed by Hugging Face and released in June 2025.

AI HardwareArtificial Intelligence

World action model

A world action model (WAM) is a robot policy design that builds action generation on a video world model backbone rather than on a vision-language model, so that a single network jointly predicts how a scene…

Embodied AIRobotics

π0.5

π0.5 (also written pi0.5, pi 0.5, or π₀.₅, and pronounced "pi zero point five") is a vision-language-action model developed by the robotics company Physical Intelligence and released on April 22, 2025.

Embodied AIRobotics

π₀ (pi-zero)

π₀ (pronounced pi-zero and sometimes written pi0 or pizero) is a vision-language-action model (VLA) developed by the robotics foundation-model startup Physical Intelligence

Embodied AIRobotics