local.ai — Charting the transition from cloud AI to local AI.
Independent benchmark quality and measured serving performance for local hardware you can buy.
Alex Cheema on Twitter / X
This is why we need open benchmarks for local AI.Otherwise it turns into tribalism and name calling.We will be publishing the largest database of open benchmarks for local AI, tested on 1,000+ real hardware setups. Every device, every interconnect, different… https://t.co/ZsU3PCdSsZ— Alex Cheema (@alexocheema) March 9, 2026
Together AI | The AI Native Cloud
Build what's next on the AI Native Cloud. Full-stack AI platform for inference, fine-tuning, and GPU clusters — powered by cutting-edge research.

12britz/awesome-free-models
A curated list of free AI models, APIs, and tools you can use without paying a cent.
Z-Space Local AI
ZAI is the Z-Space Local AI project, exploring hardware and software systems at the community level.

raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
Moltbook is the most interesting place on the internet right now
The hottest project in AI right now is Clawdbot, renamed to Moltbot, renamed to OpenClaw. It’s an open source implementation of the digital personal assistant pattern, built by Peter Steinberger …

Own your AI with local models and open protocols | Tiles Blog
A Local-First Conf talk about local models, open protocols, and user-owned AI.

Locally AI - Run AI models locally on your iPhone, iPad, and Mac.
Run Llama, Gemma, Qwen, DeepSeek, and more on your iPhone, iPad, and Mac. Optimized for Apple Silicon. Offline. Private.

AI-Native Cloud | DigitalOcean
Run AI products in production with a unified stack for agents, inference, and cloud—built for control, performance, and economics at scale.

Running local models on an M4 with 24GB memory | jola.dev
Why and How to Run Local Models in Zed

Running local models is good now
so that new Mac Studio and M5 Ultra have me thinking about 2030 now the wildcard on the $1200ish 2030 Mac Mini is RAM. who knows what the…

talat - the local transcription app for meetings, dictation, and recordings
sandsaber/Grimoire
Local Qwen isn't a worse Opus, it's a different tool

A 10 year old Xeon is all you need - point.free
ASUS Ascent GX10