







Learn which realtime and audio guide to use for each speech application.
OpenAI.fm
An interactive demo for developers to try the latest text-to-speech model in the OpenAI API

How OpenAI delivers low-latency voice AI at scale
How OpenAI rebuilt its WebRTC stack to power real-time Voice AI with low latency, global scale, and seamless conversational turn-taking.

openai/openai-realtime-agents
This is a simple demonstration of more advanced, agentic patterns built on top of the Realtime API.
DeepL AI Platform: Translation, Voice & API
Explore our AI suite and get more done: Translate speech, text, and media, or integrate the DeepL API.
Building Effective Voice Agents — Toki Sherbakov + Anoop Kotha, OpenAI
GPT-Live System Card - OpenAI Deployment Safety Hub
GPT-Live-1 and GPT-Live-1 mini are a new generation of voice models designed to make conversations with AI feel more natural and intelligent.

OpenAI's WebRTC Problem - Media over QUIC
Media over QUIC: There are ways to do voice AI without being traumatized by WebRTC.

ChatGPT Voice can keep talking while it works - Sensemaker
OpenAI’s GPT-Live splits live conversation from slower search, reasoning, and agent work in the background.
Gabber - Build Realtime AI Apps that can see, hear, and speak
Low-latency inference for VLM, TTS, and STT with orchestration for making realtime apps.

Introducing Whisper – an open source voice note taking app! Record voice notes and transcribe them into lists, blogs, & more with AI. 100% free & open source. https://t.co/UZWGkUDJ6d
Introducing Whisper – an open source voice note taking app!Record voice notes and transcribe them into lists, blogs, & more with AI.100% free & open source. pic.twitter.com/UZWGkUDJ6d— Hassan (@nutlope) July 22, 2025
Pipecat Cloud: Enterprise Voice Agents Built On Open Source - Kwindla Hultman Kramer, Daily
What makes a great ChatGPT app | OpenAI Developers
How to build capabilities that make conversations better.

Serving Voice AI at $1/hr: Open-source, LoRAs, Latency, Load Balancing - Neil Dwyer, Gabber
Our New SAM Audio Model Transforms Audio Editing
We're introducing SAM Audio, a state-of-the-art AI model that enables you to segment sound.
