







Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. $0.57 per million input tokens, $2.30 per million output tokens. 131,072 token context window, maximum output of 32,768 tokens. Includes independent benchmarks from Artificial Analysis.
Flagship Model Kimi K3 Pricing - Kimi API Platform
Kimi K3 is our flagship model for long-horizon coding and end-to-end knowledge work, with a 1M-token context window and industry-leading intelligence. The Kimi API Platform provides K3, K2.7 Code, K2.6 and other large language model APIs, supporting long context, multimodal understanding, and Tool Calling.

Kimi K3 - Kimi API Platform
Kimi K3 is our flagship model for long-horizon coding and end-to-end knowledge work, with a 1M-token context window and industry-leading intelligence. The Kimi API Platform provides K3, K2.7 Code, K2.6 and other large language model APIs, supporting long context, multimodal understanding, and Tool Calling.

Kimi K2: Open Agentic Intelligence
Kimi K2 is our latest Mixture-of-Experts model with 32 billion activated parameters and 1 trillion total parameters. It achieves state-of-the-art performance in frontier knowledge, math, and coding among non-thinking models.
On Kimi K3: Its Capabilities And Related Discontents
Kimi K3 is a very good model with excellent benchmarks.

Kimi Code - Next-Gen AI Code Agent | Automated Programming & CLI
Unlock Kimi Code (KFC), the ultimate AI toolkit for developers. Featuring high-performance CLI tools and Turbo-speed models to automate code generation and boost development efficiency. Experience faster, more reliable AI-powered coding today.
Zhuokai Zhao on Twitter / X
Tons of interesting things in the Kimi K3 tech report — here are five algorithm-side techniques that I think either I've never seen before or simply deserve more attention than they're getting.1/ They open-sourced the model but kept the speculative decoding draft model, which…— Zhuokai Zhao (@zhuokaiz) July 30, 2026
Kimi K3's weights are public. Running them is not easy. - Sensemaker
Moonshot recommends a tightly connected cluster of 64 or more AI chips.
Kimi K3: The open-weights escalation
The global implications on the AI ecosystem.

Kimi-K3/k3_tech_report.pdf at main · MoonshotAI/Kimi-K3
Open Frontier Intelligence. Contribute to MoonshotAI/Kimi-K3 development by creating an account on GitHub.
Kimi K3 Tech Blog: Open Frontier Intelligence
Kimi K3 is the world's first open 3T-class model — frontier performance across coding, knowledge work, and reasoning, with native multimodality and 1M context.
Models & Pricing | DeepSeek API Docs
The prices listed below are in units of per 1M tokens. A token, the smallest unit of text that the model recognizes, can be a word, a number, or even a punctuation mark. We will bill based on the total number of input and output tokens by the model.

The Kaitchup – AI on a Budget | Benjamin Marie | Substack
Weekly tutorials and news on adapting large language models (LLMs) to your tasks and hardware using the most recent techniques and models. The Kaitchup proposes a collection of 180+ AI notebooks regularly updated. Click to read The Kaitchup – AI on a Budget, by Benjamin Marie, a Substack publication with tens of thousands of subscribers.

5 Thoughts on Kimi K2 Thinking
Quick thoughts on another fantastic open model from a rapidly rising Chinese lab.

[New model] Kimi K3 by ZJY0516 · Pull Request #50000 · vllm-project/vllm
Purpose add moonshotai/Kimi-K3 model support Essential Elements of an Effective PR Description Checklist The purpose of the PR, such as "Fix some issue (link existing issues this PR ...
K3 is live, but its open weights aren't. Kimi's docs list 2.8T parameters, native vision and 1M context; there is no weight repo or technical report yet. For now this is a hosted-model launch. The testable open release still has to arrive. platform.kimi.ai/docs/guide/kimi-k3-quickstart