







AI powered voice to text for macOS, Windows, and iOS. Dictate in any app with offline and cloud speech recognition, 100+ languages, and custom AI modes.
VoiceInk - Advanced AI Voice Recognition for Mac
Transform Your Voice Into Text Instantly with Offline AI. Built for Mac, Optimized for Privacy. One-Time Purchase, No Subscriptions.

TypeWhisper for macOS, Windows, and iOS - Private Speech-to-Text
TypeWhisper connects private speech-to-text with system-wide dictation, files, workflows, and automation. macOS 1.6 and Windows 1.0.8 are stable; the public iOS 1.0 release is still being prepared for the App Store.

Launching a free, open-source, on-device transcription app
TL;DR – Please try Moonshine Note Taker on your Mac! For years I’ve been telling people that AI wants to be local, that on-device models aren’t just a poor man’s alternative…

DeepL AI Platform: Translation, Voice & API
Explore our AI suite and get more done: Translate speech, text, and media, or integrate the DeepL API.
Introducing Whisper – an open source voice note taking app! Record voice notes and transcribe them into lists, blogs, & more with AI. 100% free & open source. https://t.co/UZWGkUDJ6d
Introducing Whisper – an open source voice note taking app!Record voice notes and transcribe them into lists, blogs, & more with AI.100% free & open source. pic.twitter.com/UZWGkUDJ6d— Hassan (@nutlope) July 22, 2025
Voice AI & Voice Agents | An Illustrated Primer
A comprehensive guide to voice AI in 2026

MacPaw Partners with Liquid AI to Bring On-Device AI to Millions of Mac Users — Blog
MacPaw partners with Liquid AI to bring private, fast, on-device AI to millions of Mac users, starting with the Eney assistant for macOS.

Alter | Native AI built for macOS high achievers
Alter: The seamless AI that supercharges your Mac. Skip the chat, execute instant actions across all apps. 10x your productivity with complete privacy control.

Updates to Apple’s On-Device and Server Foundation Language Models
With Apple Intelligence, we're integrating powerful generative AI right into the apps and experiences people use every day, all while…

The future of Siri, or: why private inference isn’t private enough
Yesterday Apple announced a big step towards deploying real AI in their Siri ecosystem. In most ways this is good and inevitable: Siri is one of the world’s most widely-used voice agents, and…

ICML WhisperKit: On-device Real-time ASR with Billion-Scale Transformers
Real-time Automatic Speech Recognition (ASR) is a fundamental building block for many commercial applications of ML, including live captioning, dictation, meeting transcriptions, and medical scribes. Accuracy and latency are the most important factors when companies select a system to deploy. We present WhisperKit, an optimized on-device inference system for real-time ASR that significantly outperforms leading cloud-based systems. We benchmark against server-side systems that deploy a diverse set of models, including a frontier model (OpenAI gpt-4o-transcribe), a proprietary model (Deepgram nova-3), and an open-source model (Fireworks large-v3-turbo).Our results show that WhisperKit matches the lowest latency at 0.46s while achieving the highest accuracy 2.2\% WER. The optimizations behind the WhisperKit system are described in detail in this paper.
\robotoslablightdots.tts Technical Report
Text-to-speech (TTS) systems have largely solved intelligibility on standard read-speech benchmarks. What users expect from a modern system is broader: expressive and controllable output, real-time synthesis, and coverage of neutral reading, emotional dialogue, paralinguistic events, singing, and general audio. Current systems pursue this goal along three roughly distinct technical routes, and each route has its own unresolved problem.
Apple Intelligence Foundation Language Models Tech Report 2025
We introduce two multilingual, multimodal foundation language models that power Apple Intelligence features across Apple devices and…

Detail - Argmax
Detail, Apple's pick for iPad App of the Year 2025, leverages Argmax SDK to build their flagship AI features such as text-based video editing and automatic speaker switching using Argmax SDK, migrating from cloud APIs. - Dec 09, 2025
