AI Models

Explore AI Models through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: Generative AI

Articles that also belong to these categories. Counts cover all of AI Models.

Showing 1-60 of 63 articles

Amazon Nova Sonic

Amazon Nova Sonic is a real-time speech-to-speech foundation model developed by Amazon and offered through Amazon Bedrock. Announced on April 8, 2025, it is part of the Amazon Nova family of foundation models.

Generative AI

Chai-2

Chai-2 is a generative artificial intelligence model from Chai Discovery for designing antibodies and small protein binders from scratch.

AI for ScienceGenerative AI

ElevenLabs v3

Eleven v3, marketed by ElevenLabs as Eleven v3 (alpha), is a third-generation text-to-speech model that ElevenLabs released in public alpha on June 5, 2025 and described as "the most expressive Text to Speech…

Generative AISpeech & Audio AI

GAIA-4 (Wayve)

GAIA-4 is a multimodal generative world model for closed-loop autonomous driving simulation, announced by the British self-driving company Wayve on 3 August 2026 as the latest generation of its GAIA family.

Autonomous VehiclesComputer Vision

GPT Image 1

GPT Image 1 (API identifier gpt-image-1) is a natively multimodal image generation model developed by OpenAI, integrated into ChatGPT on March 25, 2025, and released as a standalone API on April 23, 2025.

Generative AIImage Generation

Gemini 3.8 Flash

Gemini 3.8 Flash is a multimodal model in Google's Gemini family. Google DeepMind released it on September 2, 2026 as a generally available model for software engineering, tool-using agents, and knowledge work.

Generative AIGoogle DeepMind

Genie 3

Genie 3 is a general-purpose foundation world model developed by Google DeepMind and announced on August 5, 2025, that generates interactive, navigable 3D environments from a single text prompt and runs in…

Generative AIGoogle DeepMind

Grok 4.5

Grok 4.5 is a proprietary multimodal large language model and reasoning model in the Grok family. It was developed by SpaceXAI in collaboration with Cursor and released through the xAI API on July 8, 2026.

Generative AILarge Language Models

Grok 4.6

Grok 4.6 is a proprietary large language model and reasoning model in the Grok family, developed by SpaceXAI and released jointly with Cursor through the xAI API on August 12, 2026.

Generative AILarge Language Models

Hedra Character

Hedra Character is a family of generative video foundation models that turn a single image plus an audio clip (and, in later versions, a text prompt) into a video in which the pictured person or character…

Generative AIVideo Generation

Hunyuan 3D

Hunyuan 3D is a family of open weight generative artificial intelligence models from Tencent that turn text prompts, single images, sketches, and other inputs into ready to use three dimensional assets…

Chinese AIGenerative AI

Imagen 3

Imagen 3 is a text-to-image generation model developed by Google DeepMind, announced at Google I/O on May 14, 2024 and progressively rolled out to users through mid-2024 and into 2025.

Generative AIGoogle

Imagen 4

Imagen 4 is the fourth-generation text-to-image model developed by Google DeepMind, announced on May 20, 2025, at Google I/O 2025.

Generative AIGoogle

Luma Dream Machine

Luma Dream Machine is a generative AI video and image platform from Luma AI (Luma Labs, Inc.), a San Francisco company, that turns text prompts and still images into short, realistic video clips.

Computer VisionGenerative AI

Lyria

Lyria is a family of AI music generation models developed by Google DeepMind, spanning text-to-music synthesis, real-time interactive music performance, and full-length song composition.

Generative AIGoogle DeepMind

MAI-Voice-1

MAI-Voice-1 is a text-to-speech (speech generation) model developed by Microsoft AI, the consumer artificial-intelligence division of Microsoft led by Mustafa Suleyman.

Generative AI

Meshy 6

Meshy 6 is the sixth major release of Meshy AI's generative 3D generation platform, a hosted generative AI service that turns text prompts or images into textured 3D models (meshes) for games, animation, and…

Generative AI

Midjourney V7

Midjourney V7 is the seventh major text-to-image generation model developed by Midjourney Inc. It launched in alpha on April 3, 2025, and became the platform's default model on June 17, 2025.

Generative AIImage Generation

NVIDIA Picasso

NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.

AI HardwareAI Inference

Nano Banana

Nano Banana is the codename, later turned official brand, for Google's native image generation and editing models built into the Gemini ecosystem and developed by Google DeepMind.

Computer VisionGenerative AI

Pika 2.5

Pika 2.5 is a generative video model developed by Pika Labs, the San Francisco based AI video startup co-founded in April 2023 by Stanford AI Lab dropouts Demi Guo and Chenlin Meng.

Computer VisionGenerative AI

Qwen-Image

Qwen-Image is an open-weight image-generation foundation model released by Alibaba's Qwen team in August 2025.

Generative AI

Qwen-Image-3.0

Qwen-Image-3.0 is a text-to-image foundation model announced by Alibaba's Qwen team on July 21, 2026, as the third generation of the Qwen-Image series .

Chinese AIGenerative AI

Recraft V3

Recraft V3 is a text-to-image generation model developed by Recraft AI and released on October 30, 2024, that became the first model to reach the number-one position on the Artificial Analysis Text-to-Image…

Generative AIImage Generation

Rodin Gen-2

Rodin Gen-2 is a generative artificial intelligence model for producing three dimensional assets from text prompts or reference images, developed by the Shanghai based company Deemos and delivered through the…

Chinese AIGenerative AI

Runway Aleph

Runway Aleph is an in-context AI video editing model from Runway that transforms and edits an existing video clip from a plain-text instruction, performing a wide range of tasks in a single model: adding…

Computer VisionGenerative AI

Seedance

Seedance is the family of foundation video generation models built by the Seed team at ByteDance, the Chinese internet company that owns TikTok and Douyin.

Chinese AIComputer Vision

Seedance 2.5

Seedance 2.5 is a proprietary AI video generation model released by ByteDance on July 31, 2026. It jointly generates audio and video from combinations of text, image, video, and audio inputs.

Chinese AIComputer Vision

Seedream

Seedream is a series of text-to-image and image-editing foundation models built by the Seed research team at ByteDance, the company behind TikTok and Douyin.

Chinese AIComputer Vision

Sesame CSM

Sesame CSM (Conversational Speech Model) is an open weights speech generation model from Sesame AI, a San Francisco startup co-founded by former Oculus chief executive Brendan Iribe.

Generative AIOpen Source AI

Shap-E

Shap-E is a conditional generative model for 3D assets developed by OpenAI, introduced in the paper "Shap-E: Generating Conditional 3D Implicit Functions" by Heewoo Jun and Alex Nichol, submitted to arXiv on…

Generative AIOpenAI

Sora 2

Sora 2 is a text-to-video and audio generation model developed by OpenAI, released on September 30, 2025, that OpenAI called "the GPT-3.5 moment for video." It succeeded the original Sora research preview from…

Generative AIMultimodal AI

Suno v5

Suno v5 is the fifth-generation AI music generation model from Suno Inc., the Cambridge, Massachusetts startup, released on September 23, 2025 to Pro and Premier subscribers as what Suno called "the world's…

Generative AIMusic & Audio Generation

Synthesia 3.0

Synthesia 3.0 is a major release of the AI video generation platform from Synthesia, the London-based company co-founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Lourdes Agapito and Matthias Niessner.

Computer VisionGenerative AI

Tripo P1

Tripo P1, marketed in full as Tripo Smart Mesh P1.0, is a production grade native 3D diffusion model that generates clean, engine ready 3D meshes from text or image prompts in as little as two seconds.

Chinese AIGenerative AI