







Kimi K3 is a very good model with excellent benchmarks.
5 Thoughts on Kimi K2 Thinking
Quick thoughts on another fantastic open model from a rapidly rising Chinese lab.

Kimi-K3/k3_tech_report.pdf at main · MoonshotAI/Kimi-K3
Open Frontier Intelligence. Contribute to MoonshotAI/Kimi-K3 development by creating an account on GitHub.
Kimi K3 Tech Blog: Open Frontier Intelligence
Kimi K3 is the world's first open 3T-class model — frontier performance across coding, knowledge work, and reasoning, with native multimodality and 1M context.
Kimi K2 0711 - API Pricing & Benchmarks
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. $0.57 per million input tokens, $2.30 per million output tokens. 131,072 token context window, maximum output of 32,768 tokens. Includes independent benchmarks from Artificial Analysis.
Zhuokai Zhao on Twitter / X
Tons of interesting things in the Kimi K3 tech report — here are five algorithm-side techniques that I think either I've never seen before or simply deserve more attention than they're getting.1/ They open-sourced the model but kept the speculative decoding draft model, which…— Zhuokai Zhao (@zhuokaiz) July 30, 2026
[New model] Kimi K3 by ZJY0516 · Pull Request #50000 · vllm-project/vllm
Purpose add moonshotai/Kimi-K3 model support Essential Elements of an Effective PR Description Checklist The purpose of the PR, such as "Fix some issue (link existing issues this PR ...
Flagship Model Kimi K3 Pricing - Kimi API Platform
Kimi K3 is our flagship model for long-horizon coding and end-to-end knowledge work, with a 1M-token context window and industry-leading intelligence. The Kimi API Platform provides K3, K2.7 Code, K2.6 and other large language model APIs, supporting long context, multimodal understanding, and Tool Calling.

Kimi K3's weights are public. Running them is not easy. - Sensemaker
Moonshot recommends a tightly connected cluster of 64 or more AI chips.
Jan on Twitter / X
Kimi K2 is **INCREDIBLE** at using tools.I built a chrome extension to chat with Google Maps, but I never posted it. All the models kept screwing up.I just replaced them with Kimi K2, and watch it easily plan an epic wine and food tour around Napa Valley for my bday!… pic.twitter.com/OOsQhVi7xL— Jan (@yawnxyz) July 15, 2025
[Kimi] Support kimi-k3 by hnyls2002 · Pull Request #32541 · sgl-project/sglang
Day-0 support for the Kimi K3 model. CI States Latest PR Test (Base): ❌ Missing run-ci label -- add it to run CI tests. Latest PR Test (Extra): ❌ Blocked -- run-ci is required first.
Kimi K3 - Kimi API Platform
Kimi K3 is our flagship model for long-horizon coding and end-to-end knowledge work, with a 1M-token context window and industry-leading intelligence. The Kimi API Platform provides K3, K2.7 Code, K2.6 and other large language model APIs, supporting long context, multimodal understanding, and Tool Calling.

Kimi K2: Open Agentic Intelligence
Kimi K2 is our latest Mixture-of-Experts model with 32 billion activated parameters and 1 trillion total parameters. It achieves state-of-the-art performance in frontier knowledge, math, and coding among non-thinking models.
Kimi K2: 1 T‑Param DeepSeek‑Inspired MoE, Re‑engineered & Trained from Scratch
Still expecting Llama 4 Behemoth? Check Kimi K2.

K3 is live, but its open weights aren't. Kimi's docs list 2.8T parameters, native vision and 1M context; there is no weight repo or technical report yet. For now this is a hosted-model launch. The testable open release still has to arrive. platform.kimi.ai/docs/guide/kimi-k3-quickstart