







ๅจ Apple Silicon Mac ไธๆฌๅฐ่ฟ่ก LLM ๆจ็ๆๅก๏ผๆไพๆฏ Ollama ๅ llama.cpp ๆดๅฟซ็ OpenAI ๅ ผๅฎน API๏ผๅๆถๅ็ๆฏๆๅทฅๅ ท่ฐ็จๅๆ็คบ็ผๅญใhttps://t.co/IXet9GW7x0Rapid-MLX ็จ Apple ่ชๅฎถ็ MLX ๆกๆถๅๆจ็๏ผๆญไบไธช FastAPI ๆๅก่ท OpenAI ๅ ผๅฎน APIใๅจ Apple Silicon ไธๆฏ Ollama ๅฟซ 2-4 ๅ๏ผ้ โฆ pic.twitter.com/d9QUbAWJSZโ Geek Lite (@QingQ77) May 3, 2026
Ollama is now powered by MLX on Apple Silicon in previewยท Ollama Blog
Today, we're previewing the fastest way to run Ollama on Apple silicon, powered by MLX, Apple's machine learning framework.

Rohan Paul on Twitter / X
๐จโ๐ง Github: Native, Apple Siliconโonly local LLM server. Similar to Ollama, but built on Apple's MLX- OpenAI API compatible, - Ollamaโcompatible- OpenAIโstyle tools + tool_choice, with tool_calls parsing and streaming deltasgithub. com/dinoki-ai/osaurus pic.twitter.com/G0NbWnEJcQโ Rohan Paul (@rohanpaul_ai) September 3, 2025

apple-silicon-llm-bench/results/complete_results.html at main ยท AlexHiesch/apple-silicon-llm-bench
Systematic LLM inference benchmark for Apple Silicon: 8 backends, 7 models, 791 measurements - AlexHiesch/apple-silicon-llm-bench
Silicon Jungle on Twitter / X
iโm glad theyโre doing what theyโre doing.but they have one thing wrong.this should be happening at the OS level, not the browser. https://t.co/GiSLenJ0bgโ Silicon Jungle (@JungleSilicon) December 2, 2024
mzau/broke-cluster
A Poor Man's Apple Silicon LLM Cluster โ tuned for MLX, scalable without shame.
Cua on Twitter / X
1/ Today, as part of our broader research into Apple Silicon virtualization, we're releasing a process-scoped Metal capability layer for macOS VMs. On one M1 Ultra, prompt / generation:TinyLlama: 11.08ร / 16.36รGemma 4 12B: 7.20ร / 14.54รMuse Glimmer 30B: 7.55ร / 8.87ร pic.twitter.com/6UwxgoDowTโ Cua (@trycua) August 11, 2026

How Apple Can Own I/O to Own the Universe
Plus! The Great Inflation; Two-Way API Businesses; Zoom and Cities; From ETrade to eToro; Hardware-as-a-Service

Part 4: Brief history of Apple ML Stack
By Mirai Labs, frontier on-device AI lab. Building the models, inference runtime, and quantization stack from the device constraint up.

Protected by its moat, Apple has time to get AI right
Best of all, the hardware it sells today will run whatever Apple comes up with.

็จMacbookๅพฎ่ฐQwen3๏ผๆๆๆๆไฝ ็จๅพฎ่ฐ็ปQwen่ตทไธไธชๆฐๅๅญ
clandestine.eth ๐ฆ๐ on Twitter / X
Heterogeneous acceleration on Apple Silicon achieved.ANE + GPU running in parallel.Mirror SD with DFlash, ported to MLX โ targeting ANE + GPU simultaneously.The M-series was designed for this. We just hadn't unlocked it yet. pic.twitter.com/raSH0CMN4Vโ clandestine.eth ๐ฆ๐ (@0xClandestine) April 15, 2026

Latest News - Apple Developer
Learn about the latest technologies, events, and policies for developers.

๐ค ๐ด๐๐๐๐ ๐น๐๐ข๐๐๐๐ก๐๐๐ ๐๐๐๐๐๐ - ๐ถ๐๐๐ข๐๐ ๐ด๐๐ผ ๐ท๐๐๐ by Anthropic Plug Claude into Apple's Foundation Models framework. Swap between on-device and frontier models behind the same ๐ฟ๐๐๐๐ข๐๐๐๐๐๐๐๐๐๐๐ ๐ ๐๐๐ API. #Swift #AI #FoundationModels platform.claude.com/docs/en/cli-sdks-libraries/liโฆ
Apple Foundation Models
platform.claude.com