







Moonshot recommends a tightly connected cluster of 64 or more AI chips.
Kimi-K3/k3_tech_report.pdf at main · MoonshotAI/Kimi-K3
Open Frontier Intelligence. Contribute to MoonshotAI/Kimi-K3 development by creating an account on GitHub.
On Kimi K3: Its Capabilities And Related Discontents
Kimi K3 is a very good model with excellent benchmarks.

Kimi K3: The open-weights escalation
The global implications on the AI ecosystem.

Kimi K2 0711 - API Pricing & Benchmarks
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. $0.57 per million input tokens, $2.30 per million output tokens. 131,072 token context window, maximum output of 32,768 tokens. Includes independent benchmarks from Artificial Analysis.
Open-weight AI is having its Kubernetes moment. Let's not ruin it. | Tobi Knaup
Open-weight models are becoming the foundation for the next AI ecosystem. The US should compete in it, not wall itself off.

Fastino trains AI models on cheap gaming GPUs and just raised $17.5M led by Khosla | TechCrunch
Tech giants like to boast about trillion-parameter AI models that require massive and expensive GPU clusters. But Fastino is taking a different approach.

moonshotai/Kimi-K3 at 2496450e92e425c886db095102a52a6682ca3970
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
moonshotai/Kimi-K3 · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Post-Training 50x Faster
We're announcing Trellis, the fastest open-source post-training code for Kimi K2 Thinking

Open Weight AI Models Explained for Everyone
This might be bigger than DeepSeek
Alex Cheema on Twitter / X
It’s kind of crazy but the shitstorm of supply chain issues has created a new best-in-class local AI deployment: M5 Max MacBook clusters.- The memory unit economics are great - each MacBook has 128GB @ 614GB/s for $5k- M5 Max added tensor cores (Apple Neural Accelerators) with… https://t.co/f8STQ0tLZs pic.twitter.com/FLm3oOnyEl— Alex Cheema (@alexocheema) May 14, 2026

Kimi K3 Tech Blog: Open Frontier Intelligence
Kimi K3 is the world's first open 3T-class model — frontier performance across coding, knowledge work, and reasoning, with native multimodality and 1M context.
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
K3 is live, but its open weights aren't. Kimi's docs list 2.8T parameters, native vision and 1M context; there is no weight repo or technical report yet. For now this is a hosted-model launch. The testable open release still has to arrive. platform.kimi.ai/docs/guide/kimi-k3-quickstart