







Experience natural speech synthesis with Coqui TTS and XTTS V2 technology. Features voice cloning and support for 8 languages.
\robotoslablightdots.tts Technical Report
Text-to-speech (TTS) systems have largely solved intelligibility on standard read-speech benchmarks. What users expect from a modern system is broader: expressive and controllable output, real-time synthesis, and coverage of neutral reading, emotional dialogue, paralinguistic events, singing, and general audio. Current systems pursue this goal along three roughly distinct technical routes, and each route has its own unresolved problem.
make ai speak computer by dottxt @ Nouscon 2024
Announcing transcribe.cpp
Meet transcribe.cpp, a new open-source C/C++ speech-to-text inference library with portable, GPU-accelerated support for multiple STT models. Developed through Mozilla.ai's Builders in Residence program, it makes adding fast, local transcription to applications easier than ever.

OpenAI.fm
An interactive demo for developers to try the latest text-to-speech model in the OpenAI API

Introducing GPT-Live
A new generation of voice models for natural human-AI interaction, now powering ChatGPT Voice.

OpenAI's WebRTC Problem - Media over QUIC
Media over QUIC: There are ways to do voice AI without being traumatized by WebRTC.

DeepL AI Platform: Translation, Voice & API
Explore our AI suite and get more done: Translate speech, text, and media, or integrate the DeepL API.
Signal creator Moxie Marlinspike wants to do for AI what he did for messaging
Introducing Confer, an end-to-end AI assistant that just works.

Voice AI & Voice Agents | An Illustrated Primer
A comprehensive guide to voice AI in 2026

Beautiful UI — Crafted primitives for AI-native interfaces
A small library of extremely crafted, copy-paste components for chat agents, thinking states, human-in-the-loop approvals, and everything agents need to talk to humans beautifully.
Introducing Whisper – an open source voice note taking app! Record voice notes and transcribe them into lists, blogs, & more with AI. 100% free & open source. https://t.co/UZWGkUDJ6d
Introducing Whisper – an open source voice note taking app!Record voice notes and transcribe them into lists, blogs, & more with AI.100% free & open source. pic.twitter.com/UZWGkUDJ6d— Hassan (@nutlope) July 22, 2025
Introducing **transcribe.cpp** 🎙️ A new open-source C/C++ speech-to-text inference library for fast, local transcription. ✅ Multiple GGUF STT models ✅ GPU acceleration (Metal, Vulkan & CUDA) ✅ Portable across platforms Built through @mozilla.ai's BiR program. Blog: blog.mozilla.ai/announcing-transcribe-cpp/
Announcing transcribe.cpp
blog.mozilla.aiI think of all of the AI / ML / CS tech out there, speech generation freaks me out the most.
Opensourcing TADA: Fast, Reliable Speech Generation Through Text-Acoustic Synchronization
www.hume.ai