GPT-Live
GPT-Live is a family of full-duplex voice models developed by OpenAI for continuous spoken interaction.
Explore OpenAI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of OpenAI.
Showing 1-6 of 6 articles
GPT-Live is a family of full-duplex voice models developed by OpenAI for continuous spoken interaction.
GPT-Realtime is a family of speech-to-speech models exposed through the OpenAI Realtime API, a low-latency interface that lets developers build voice agents which take spoken audio in and return spoken audio…
GPT-Transcribe (gpt-transcribe) and GPT-Live-Transcribe (gpt-live-transcribe) are two closed-weight speech-to-text models that OpenAI added to its API on July 28, 2026.
The OpenAI Realtime API is a speech-to-speech interface from OpenAI that lets developers build low-latency, bidirectional voice AI applications powered by GPT-4o and later by the purpose-built gpt-realtime…
Voice Engine is a speech-generation and voice-cloning model developed by OpenAI that can produce natural-sounding speech resembling a specific person from a single audio sample as short as 15 seconds.
Whisper is an open-source family of automatic speech recognition (ASR) models developed by OpenAI and first released on September 21, 2022.