







Liquid Audio - Speech-to-Speech audio models by Liquid AI
Liquid AI — Device-native foundation models.
Liquid AI is an efficiency-first foundation model company. We build highly capable, compute-optimized models that bring intelligence to any device and medium of choice.

Our New SAM Audio Model Transforms Audio Editing
We're introducing SAM Audio, a state-of-the-art AI model that enables you to segment sound.

Voice AI & Voice Agents | An Illustrated Primer
A comprehensive guide to voice AI in 2026

Liquid AI on Twitter / X
Today, we release LFM2.5-350M. Agentic loops at 350M parameters.A 350M model trained for reliable data extraction and tool use, where models at this scale typically struggle.<500MB when quantized, built for environments where compute, memory, and latency are constrained.🧵 pic.twitter.com/zZPKzcCwH9— Liquid AI (@liquidai) March 31, 2026

Liquid AI Launches LEAP and Liquid Apollo: The Easiest Way to Build with On-Device AI | Liquid AI
Today marks a pivotal milestone in the evolution of edge AI. Liquid AI is thrilled to announce LEAP v0, our first developer-ready platform for on-device AI deployment—and Liquid Apollo, a lightweight iOS-native application built to showcase and stress-test small foundation models directly on your phone.

Liquid <> .txt Collaboration | Liquid AI
Faster and more accurate function calling on the edge with .txt’s structured outputs and Liquid Foundation Models

Stability AI optimized its audio generation model to run on Arm chips | TechCrunch
AI startup Stability AI has teamed up with chipmaker Arm to bring Stability's Stable Audio Open, an AI model that can generate audio including sound

Introducing LFM2: The Fastest On-Device Foundation Models on the Market | Liquid AI
Today, we release LFM2, a new class of Liquid Foundation Models (LFMs) that sets a new standard in quality, speed, and memory efficiency for on-device deployment. Built on a hybrid architecture, LFM2 delivers 200% faster decode and prefill performance than Qwen3 and Gemma 3 on CPU. It also significantly outperforms models in each size class on instruction-following and function calling—the core capabilities that make LLMs reliable for building AI agents.

Automotive — Solutions — Liquid AI
On-device AI for automakers — real-time, personalized in-car assistants that run on the vehicle's existing CPUs and NPUs.

DeepL AI Platform: Translation, Voice & API
Explore our AI suite and get more done: Translate speech, text, and media, or integrate the DeepL API.
\robotoslablightdots.tts Technical Report
Text-to-speech (TTS) systems have largely solved intelligibility on standard read-speech benchmarks. What users expect from a modern system is broader: expressive and controllable output, real-time synthesis, and coverage of neutral reading, emotional dialogue, paralinguistic events, singing, and general audio. Current systems pursue this goal along three roughly distinct technical routes, and each route has its own unresolved problem.
OpenAI.fm
An interactive demo for developers to try the latest text-to-speech model in the OpenAI API

Introducing GPT-Live
A new generation of voice models for natural human-AI interaction, now powering ChatGPT Voice.

Coqui TTS & XTTS V2: AI Text to Speech in 8 Languages
Experience natural speech synthesis with Coqui TTS and XTTS V2 technology. Features voice cloning and support for 8 languages.
Realtime and audio | OpenAI API
Learn which realtime and audio guide to use for each speech application.

I think of all of the AI / ML / CS tech out there, speech generation freaks me out the most.
Opensourcing TADA: Fast, Reliable Speech Generation Through Text-Acoustic Synchronization
www.hume.ai