







Mlem is a beautiful, intuitive open source iOS client for Lemmy that lets you effortlessly participate in conversations across all Lemmy servers.
Rohan Paul on Twitter / X
👨🔧 Github: Native, Apple Silicon–only local LLM server. Similar to Ollama, but built on Apple's MLX- OpenAI API compatible, - Ollama‑compatible- OpenAI‑style tools + tool_choice, with tool_calls parsing and streaming deltasgithub. com/dinoki-ai/osaurus pic.twitter.com/G0NbWnEJcQ— Rohan Paul (@rohanpaul_ai) September 3, 2025

google-ai-edge/gallery
A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.
Interstellar - Apps on Google Play
An open source Mbin/Lemmy/PieFed client, connecting you to the fediverse.
Apple stumbled into succes with MLX
201 votes, 76 comments. Qwen3-next 80b-a3b is out in mlx on hugging face, MLX already supports it. Open source contributors got this done within 2…
Google for Developers Blog - News about Web, Mobile, AI and Cloud
LiteRT is the universal framework for on-device AI. The production stack delivers 1.4x faster cross-platform GPU performance, streamlined NPU acceleration, and superior GenAI support for open models like Gemma.

Alex Albert on Twitter / X
Introducing the Model Context Protocol (MCP)An open standard we've been working on at Anthropic that solves a core challenge with LLM apps - connecting them to your data.No more building custom integrations for every data source. MCP provides one protocol to connect them all: pic.twitter.com/kYsivQyPDq— Alex Albert (@alexalbert__) November 25, 2024

google-ai-edge/LiteRT
LiteRT, successor to TensorFlow Lite. is Google's On-device framework for high-performance ML & GenAI deployment on edge platforms, via efficient conversion, runtime, and optimization
Part 4: Brief history of Apple ML Stack
By Mirai Labs, frontier on-device AI lab. Building the models, inference runtime, and quantization stack from the device constraint up.

Unsloth AI on Twitter / X
Introducing Unsloth Desktop 🦥The first desktop app to run and train models locally.• Open-source. Runs on Mac, Windows and Linux• Supports MLX, diffusion image/video, audio, GGUF• Connect Claude Code and Codex to local LLMs• 50% more accurate, self-healing tool calls +… pic.twitter.com/vjTFB1e5IQ— Unsloth AI (@UnslothAI) August 11, 2026
I don't pay for ChatGPT, Perplexity, Gemini, or Claude – I stick to my self-hosted LLMs instead
There's no point in relying on AI tools when my local LLMs can handle everything

elements
As much as I struggle with on-device processing and the quality of its output compared to server models, I am excited by some of the APIs that are being built into browsers that are backed by LLMs and other AI inference models. For example, the prompt API, along with a multi-modal version that can take any arbitrary combination of text, image, and audio and run prompts against them. These APIs are neat but not yet web-exposed and many developers struggle to know what to do with a generic prompt. It’s not a solution that is natural to many people yet.

Atomic Chat: Free Local AI Chat for Mac, Windows & iPhone | Private & Offline
Free, open-source local AI chat for Mac, Windows & iPhone. Run Gemma, Qwen, DeepSeek, Llama offline. 1,000+ models, no cloud, no subscription. Download free.

Locally AI - Run AI models locally on your iPhone, iPad, and Mac.
Run Llama, Gemma, Qwen, DeepSeek, and more on your iPhone, iPad, and Mac. Optimized for Apple Silicon. Offline. Private.

interstellar-app/interstellar
An app for Mbin/Lemmy/PieFed, connecting you to the fediverse.
MLX India Community Meetup 2 | Own Your AI by Ankesh Bharti