







SoTA open-source TTS
Coqui TTS & XTTS V2: AI Text to Speech in 8 Languages
Experience natural speech synthesis with Coqui TTS and XTTS V2 technology. Features voice cloning and support for 8 languages.
make ai speak computer by dottxt @ Nouscon 2024
\robotoslablightdots.tts Technical Report
Text-to-speech (TTS) systems have largely solved intelligibility on standard read-speech benchmarks. What users expect from a modern system is broader: expressive and controllable output, real-time synthesis, and coverage of neutral reading, emotional dialogue, paralinguistic events, singing, and general audio. Current systems pursue this goal along three roughly distinct technical routes, and each route has its own unresolved problem.
LibreChat - The Open-Source AI Platform
LibreChat brings together all your AI conversations in one unified, customizable interface.
Launching a free, open-source, on-device transcription app
TL;DR – Please try Moonshine Note Taker on your Mac! For years I’ve been telling people that AI wants to be local, that on-device models aren’t just a poor man’s alternative…

LibreChat | Proxmox VE Helper Scripts
LibreChat is an open-source AI chat platform that supports multiple AI providers including OpenAI, Anthropic, Google, and more. It features conversation…
Search results | Cambridge Network
TTP exists to innovate. We create, test and develop new technologies and products, that change industries for the better. We challenge and are challenged. Every day. We're the scientists, engineers and designers re-inventing the way millions of people live and work.
Gabber - Build Realtime AI Apps that can see, hear, and speak
Low-latency inference for VLM, TTS, and STT with orchestration for making realtime apps.

Beautiful UI — Crafted primitives for AI-native interfaces
A small library of extremely crafted, copy-paste components for chat agents, thinking states, human-in-the-loop approvals, and everything agents need to talk to humans beautifully.
browseros-ai/BrowserOS
🌐 The open-source Agentic browser; alternative to ChatGPT Atlas, Perplexity Comet, Dia.
Serving Voice AI at $1/hr: Open-source, LoRAs, Latency, Load Balancing - Neil Dwyer, Gabber
AI code and software craft - alex wennerberg
Much has been said about audio, video and text "slop": low-quality, AI-generated content that has proliferated on the internet since the release of publicly-accessible AI models. Garbage content has always existed online, but the novelty of AI is that it has made its generation orders of magnitude less labor-intensive. For anyone who lacks a discerning eye, or is doing some task where discernment simply does not matter, AI has become a sufficient replacement for human hands.
Signal creator Moxie Marlinspike wants to do for AI what he did for messaging
Introducing Confer, an end-to-end AI assistant that just works.

OpenAI.fm
An interactive demo for developers to try the latest text-to-speech model in the OpenAI API

elements
As much as I struggle with on-device processing and the quality of its output compared to server models, I am excited by some of the APIs that are being built into browsers that are backed by LLMs and other AI inference models. For example, the prompt API, along with a multi-modal version that can take any arbitrary combination of text, image, and audio and run prompts against them. These APIs are neat but not yet web-exposed and many developers struggle to know what to do with a generic prompt. It’s not a solution that is natural to many people yet.
