







Run Llama, Gemma, Qwen, DeepSeek, and more on your iPhone, iPad, and Mac. Optimized for Apple Silicon. Offline. Private.
Atomic Chat: Free Local AI Chat for Mac, Windows & iPhone | Private & Offline
Free, open-source local AI chat for Mac, Windows & iPhone. Run Gemma, Qwen, DeepSeek, Llama offline. 1,000+ models, no cloud, no subscription. Download free.

Osaurus — Own Your AI on Apple Silicon
Own your AI: local-first agents with memory, tools, and identity on Apple Silicon. Offline, open source, and API-compatible with OpenAI, Anthropic, and Ollama.

Osaurus — Own Your AI on Apple Silicon
Own your AI: local-first agents with memory, tools, and identity on Apple Silicon. Offline, open source, and API-compatible with OpenAI, Anthropic, and Ollama.

Apple will reportedly open up its local AI models to third-party apps
Use the models Apple Intelligence uses.

AI-Native Cloud | DigitalOcean
Run AI products in production with a unified stack for agents, inference, and cloud—built for control, performance, and economics at scale.
Melange | On-device AI for Mobile
Select. Benchmark. Deploy | End-to-end on-device AI deployment tool for mobile devs.
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
mudler/LocalAI
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
Together AI | The AI Native Cloud
Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research.

Mirai Labs: Frontier On-Device AI Lab
Models, runtime & infrastructure to make on-device AI interactive, ambient & continuous.

Mount Thor — AI Execution Environments on Apple Hardware
Managed macOS environments for AI workloads that require native desktop access, persistent state, or model inference on Apple silicon.

Alex Cheema on Twitter / X
This is why we need open benchmarks for local AI.Otherwise it turns into tribalism and name calling.We will be publishing the largest database of open benchmarks for local AI, tested on 1,000+ real hardware setups. Every device, every interconnect, different… https://t.co/ZsU3PCdSsZ— Alex Cheema (@alexocheema) March 9, 2026
Lilypad-Tech/lilypad
Run AI workloads easily in a decentralized GPU network. https://www.youtube.com/watch?v=yQnB2Yxia4Y
Open Minis — Your Private On-Device AI Agent
Open Minis is an on-device AI agent for iPhone, iPad, Mac, Vision Pro and Android. Browse the web, manage your schedule, control smart home, and automate tasks — all privately on your device.
Part 4: Brief history of Apple ML Stack
By Mirai Labs, frontier on-device AI lab. Building the models, inference runtime, and quantization stack from the device constraint up.
