Speech & Audio AI

Explore Speech & Audio AI through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: OpenAI

Articles that also belong to these categories. Counts cover all of Speech & Audio AI.

Showing 1-6 of 6 articles

GPT-Live

GPT-Live is a family of full-duplex voice models developed by OpenAI for continuous spoken interaction.

AI ModelsOpenAI

GPT-Realtime / OpenAI Realtime API

GPT-Realtime is a family of speech-to-speech models exposed through the OpenAI Realtime API, a low-latency interface that lets developers build voice agents which take spoken audio in and return spoken audio…

OpenAIVoice AI

GPT-Transcribe

GPT-Transcribe (gpt-transcribe) and GPT-Live-Transcribe (gpt-live-transcribe) are two closed-weight speech-to-text models that OpenAI added to its API on July 28, 2026.

AI ModelsOpenAI

OpenAI Realtime API

The OpenAI Realtime API is a speech-to-speech interface from OpenAI that lets developers build low-latency, bidirectional voice AI applications powered by GPT-4o and later by the purpose-built gpt-realtime…

Conversational AIDeveloper Tools

Voice Engine (OpenAI)

Voice Engine is a speech-generation and voice-cloning model developed by OpenAI that can produce natural-sounding speech resembling a specific person from a single audio sample as short as 15 seconds.

OpenAIVoice AI