Lyria
Lyria is a family of AI music generation models developed by Google DeepMind, spanning text-to-music synthesis, real-time interactive music performance, and full-length song composition.
Explore AI Models through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Models.
Showing 241-300 of 408 articles
Lyria is a family of AI music generation models developed by Google DeepMind, spanning text-to-music synthesis, real-time interactive music performance, and full-length song composition.
MAGI-2 Preview is a public research release of a unified audio-video generation model developed by Sand.ai.
MAI-1-preview is a large language model developed by Microsoft AI, the consumer artificial intelligence division of Microsoft led by Mustafa Suleyman.
MAI-Code-1 is the name commonly used for Microsoft's first in-house coding model, which the company shipped as MAI-Code-1-Flash at the Microsoft Build 2026 developer conference on June 2, 2026.
MAI-Image-2.6 is a text-to-image and image-editing model built in-house by Microsoft AI, the division led by Mustafa Suleyman.
MAI-Voice-1 is a text-to-speech (speech generation) model developed by Microsoft AI, the consumer artificial-intelligence division of Microsoft led by Mustafa Suleyman.
Mamba 2 is a state space model architecture introduced in the paper "Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality" by Tri Dao and Albert Gu
MatterGen is a generative artificial intelligence model from Microsoft Research for designing novel inorganic crystalline materials.
Meshy 6 is the sixth major release of Meshy AI's generative 3D generation platform, a hosted generative AI service that turns text prompts or images into textured 3D models (meshes) for games, animation, and…
Meta Motivo is a behavioral foundation model for controlling a simulated humanoid body, released by Meta AI's Fundamental AI Research (FAIR) group on December 12, 2024.
Microsoft MAI is the family of first-party artificial intelligence models built in-house by Microsoft AI, the consumer AI division that Microsoft created in March 2024 and put under Mustafa Suleyman.
Midjourney V7 is the seventh major text-to-image generation model developed by Midjourney Inc. It launched in alpha on April 3, 2025, and became the platform's default model on June 17, 2025.
MiniCPM is a family of compact, openly licensed language models published by OpenBMB, the shared open-source brand of Tsinghua University's natural language processing lab (THUNLP) and the Beijing company…
MiniCPM5-2B is an open-weights small language model published by OpenBMB on September 7, 2026 under the Apache License 2.0.
MiniMax H3 (also written MiniMax-H3) is a multimodal video generation model developed by MiniMax, the Shanghai-based AI company that operates the Hailuo AI video service.
MiniMax M2 is an open-weight large language model released on October 27, 2025 by the Shanghai-based AI company MiniMax, built as a Mixture of Experts model with 230 billion total parameters and roughly 10…
Ministral is a family of two small language models released by the French artificial intelligence company Mistral AI on October 16, 2024.
Mistral Large is the family of flagship large language models developed by Mistral AI, the Paris-based AI laboratory founded in 2023, and is the company's most capable general-purpose model line.
Mistral Large 3 is a sparse mixture-of-experts large language model released on December 2, 2025 by the French AI company Mistral AI, distributed as open weights under the Apache 2.0 license with roughly 675…
Mistral Medium 3 is a proprietary multimodal large language model developed by Mistral AI and released on May 7, 2025.
Mistral OCR 3 is a document-understanding and optical character recognition model from Mistral AI, released in mid-December 2025 as the third generation of the company's OCR product line.
Mistral OCR 4 is a proprietary document extraction model and service developed by Mistral AI. Released on June 23, 2026, it converts documents and page images into Markdown and structured layout data.
Mixtral 8x22B is a sparse mixture-of-experts (MoE) large language model released by the French AI company Mistral AI on April 17, 2024.
Molmo is a family of open-weight, open-data vision-language models (VLMs) released by the Allen Institute for AI (Ai2) on 25 September 2024.
Moshi is a full-duplex speech-to-speech foundation model developed by Kyutai, a French nonprofit artificial intelligence research laboratory.
MuZero is a model-based reinforcement learning algorithm developed by DeepMind that masters Go, chess, shogi, and 57 Atari video games at superhuman or state of the art level without ever being told the rules…
Muse Glimmer is an open-weight text-and-image model developed by Meta AI for local agent and coding workloads. Meta released the model on August 10, 2026 under the identifier meta-models/Muse-Glimmer-30B.
Muse Image is a proprietary image-generation and editing system developed by Meta Superintelligence Labs, a division of Meta.
Muse Voice Transcribe is a hosted speech recognition model developed by Meta Superintelligence Labs. Meta released it on September 1, 2026 as the lab's first real-time audio perception model.
NVIDIA Alpamayo 2 Super is an open, 34-billion-parameter reasoning-based vision-language-action model (VLA) for safe, Level 4 robotaxi and autonomous-vehicle development from NVIDIA.
NVIDIA COMPASS is a framework and trained model for cross-embodiment robot navigation. Its name expands to Cross-Embodiment Mobility Policy via Residual RL and Skill Synthesis.
Canary is a family of open speech models developed by Nvidia as part of its NeMo conversational AI toolkit.
NVIDIA Cosmos is a world foundation model platform developed by NVIDIA for physical AI applications, including autonomous vehicles and robotics.
NVIDIA Earth-2 is a software platform and family of models from NVIDIA for AI-accelerated weather and climate simulation, prediction, and visualization.
NVIDIA Isaac GR00T N1 is an open foundation model for humanoid robots developed by NVIDIA and unveiled by Jensen Huang on March 18, 2025 at the company's annual GTC conference in San Jose, California.
NVIDIA Isaac GR00T N1.6 is a 3-billion-parameter vision-language-action model and cross-embodiment robot foundation model developed by Nvidia.
NVIDIA Isaac GR00T N1.7 is a 3-billion-parameter, cross-embodiment vision-language-action model developed by NVIDIA for humanoid and manipulation robots.
NVIDIA Ising is a family of open AI models from NVIDIA for operating quantum computers, covering two of the field's main engineering bottlenecks: quantum processor calibration and quantum error-correction…
NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.
Nano Banana is the codename, later turned official brand, for Google's native image generation and editing models built into the Gemini ecosystem and developed by Google DeepMind.
Nano Banana 2 Lite is a text-to-image generation and image editing model released by Google DeepMind on June 30, 2026.
Nemotron is NVIDIA's brand for its family of open large language models and the datasets, training recipes, and evaluation tools built around them.
Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…
NVIDIA Nemotron 3.5 Lightning is an open-weights 30 billion parameter mixture-of-experts language model with 3 billion active parameters per token, released by NVIDIA on August 11
Nemotron Nano 2 is a family of small, open-weight reasoning language models released by NVIDIA on August 18, 2025
Nemotron-4 is a family of decoder-only large language models developed by NVIDIA and documented in two technical reports released in 2024.
Nemotron-H is a family of open-weight large language models released by NVIDIA in April 2025 that replace most of the self-attention layers of a standard Transformer with Mamba-2 state-space layers, producing…
Nemotron-Labs-TwoTower is an open-weight diffusion language model released by NVIDIA in mid-2026.
North Mini Code is an open-weight large language model developed by Cohere for agentic software development and AI code generation.
Nougat (Neural Optical Understanding for Academic Documents) is a document-understanding model from Meta AI that converts the rendered image of a document page into structured markup text.
OLMo (Open Language Model) is a family of fully open large language models built by the Allen Institute for AI (Ai2) and first released on February 1, 2024.
OLMo 2 is the second generation of fully open large language models released by the Allen Institute for AI (Ai2), spanning 7B, 13B, and 32B parameter sizes.
OLMo 3 is the third generation of fully open language models released by the Allen Institute for AI (Ai2).
OLMoE (Open Mixture-of-Experts) is a fully open sparse mixture of experts large language model released by the Allen Institute for AI (Ai2) on September 3, 2024 .
Open weights is the practice of publishing the trained parameter tensors of a neural network so that anyone can download, run, fine-tune, and redistribute the model without obtaining a separate inference…
OpenAI Codex is the brand OpenAI has used for two distinct generations of code-focused artificial intelligence products: a 2021 large language model that turned natural-language prompts into code
OpenAI o1 is a family of proprietary large language models developed by OpenAI and trained to use additional computation before returning an answer.
OpenAI o3 is a family of reasoning-focused large language models developed by OpenAI and the second generation of the company's o-series reasoning models, best known for scoring 87.5% on the ARC-AGI…
OpenPI (stylized openpi) is the open-source repository of robot foundation models, training code, and inference utilities published by Physical Intelligence, the San Francisco robotics and AI startup…
OpenVLA is a 7-billion-parameter open-source vision-language-action model (VLA) for robotic manipulation, released in June 2024 by a collaboration of researchers from Stanford University, UC Berkeley, the…