AudioLM
AudioLM is a framework from Google Research for generating high-quality audio by treating the problem as a language-modeling task over discrete tokens.
Explore Google through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Google.
Showing 1-4 of 4 articles
AudioLM is a framework from Google Research for generating high-quality audio by treating the problem as a language-modeling task over discrete tokens.
DolphinGemma is an audio language model developed by Google to help scientists analyze the vocalizations of wild dolphins.
Gemini 3.5 Transcribe is a family of speech recognition models developed by Google DeepMind and introduced by Google on August 26, 2026.
SoundStream is an end-to-end neural audio codec introduced by Google Research in July 2021 that compresses speech, music, and general audio at low-to-medium bitrates ranging from 3 kbps to 18 kbps on 24 kHz…