Audio-to-Audio Models
Audio-to-audio models are machine learning systems that take an audio waveform as input and produce a different audio waveform as output.
Explore Music & Audio Generation through related topics and the articles other pages reference most.
Ranked by links from other AI Wiki pages.
Articles that also belong to these categories. Counts cover all of Music & Audio Generation.
Showing 1-23 of 23 articles
Audio-to-audio models are machine learning systems that take an audio waveform as input and produce a different audio waveform as output.
AudioLM is a framework from Google Research for generating high-quality audio by treating the problem as a language-modeling task over discrete tokens.
Audiobox is a foundation research model for audio generation developed by Meta AI and its Fundamental AI Research (FAIR) group.
As of July 2026, the best AI music generator overall is Suno: its v5.5 model produces the most natural vocals, the deepest editing toolkit (Suno Studio, an AI-native DAW with stem separation and MIDI export)…
Boomy is a generative artificial intelligence music platform that lets people without musical training assemble original songs in a web browser and publish them, under their own artist name, to streaming…
ElevenLabs Music (also marketed as Eleven Music) is an AI music generation product developed by ElevenLabs, the voice AI company founded in 2022 by Piotr Dabkowski and Mati Staniszewski.
Jukebox is a neural network for music generation developed by OpenAI that produces music, including rudimentary singing, as raw audio across a range of genres and artist styles.
Lyria is a family of AI music generation models developed by Google DeepMind, spanning text-to-music synthesis, real-time interactive music performance, and full-length song composition.
Lyria 2 is a high-fidelity, text-to-music generation model built by Google DeepMind that turns text prompts into professional-grade instrumental audio.
Lyria 3.5 is a music generation model from Google DeepMind and the newest member of the Lyria family.
Magenta is an open-source research project from Google that explores the role of machine learning in creating art and music.
MuseNet is a deep neural network for symbolic music generation, announced by OpenAI on April 25, 2019.
Music ChatGPT plugins were a category of third-party extensions for ChatGPT that let the chatbot generate streaming playlists, recommend and surface songs, convert music notation to audio and sheet music…
MusicGen is an open-weights text-to-music generation model from Meta AI's Fundamental AI Research (FAIR) team, released on 8 June 2023, that generates roughly 30-second music clips from a text prompt, a…
MusicLM is a text-to-music generation model from Google Research that generates high-fidelity music at 24 kHz from natural language descriptions and keeps that audio consistent over several minutes.
Sonauto is a generative artificial intelligence music platform that converts text prompts, lyrics, and melody inputs into complete songs with vocals and instrumentation.
Stable Audio is a family of generative AI models from Stability AI that turn a text prompt into music or sound effects as a stereo audio file.
Stable Audio 2.5 is an enterprise focused text-to-audio generation model released by Stability AI on September 10, 2025.
Suno is a generative artificial intelligence company that develops a text-to-music platform capable of producing complete songs, including vocals, instrumentals, and lyrics, from simple text prompts.
Suno v5 is the fifth-generation AI music generation model from Suno Inc., the Cambridge, Massachusetts startup, released on September 23, 2025 to Pro and Premier subscribers as what Suno called "the world's…
UMG Recordings, Inc., et al. v. Suno, Inc. is a landmark copyright infringement lawsuit filed on June 24, 2024 by a coalition of major record labels
UMG Recordings, Inc., et al. v. Uncharted Labs, Inc. (case number 1:24-cv-04777) is a federal copyright infringement lawsuit filed on June 24
Udio is an artificial intelligence music generation platform developed by Uncharted Labs, Inc. that creates full songs, complete with vocals, instrumentation, and lyrics, from a single text prompt.