







Day-0 support for the Kimi K3 model. CI States Latest PR Test (Base): ❌ Missing run-ci label -- add it to run CI tests. Latest PR Test (Extra): ❌ Blocked -- run-ci is required first.
[New model] Kimi K3 by ZJY0516 · Pull Request #50000 · vllm-project/vllm
Purpose add moonshotai/Kimi-K3 model support Essential Elements of an Effective PR Description Checklist The purpose of the PR, such as "Fix some issue (link existing issues this PR ...
Kimi-K3/k3_tech_report.pdf at main · MoonshotAI/Kimi-K3
Open Frontier Intelligence. Contribute to MoonshotAI/Kimi-K3 development by creating an account on GitHub.
On Kimi K3: Its Capabilities And Related Discontents
Kimi K3 is a very good model with excellent benchmarks.

Flagship Model Kimi K3 Pricing - Kimi API Platform
Kimi K3 is our flagship model for long-horizon coding and end-to-end knowledge work, with a 1M-token context window and industry-leading intelligence. The Kimi API Platform provides K3, K2.7 Code, K2.6 and other large language model APIs, supporting long context, multimodal understanding, and Tool Calling.

Kimi K3 - Kimi API Platform
Kimi K3 is our flagship model for long-horizon coding and end-to-end knowledge work, with a 1M-token context window and industry-leading intelligence. The Kimi API Platform provides K3, K2.7 Code, K2.6 and other large language model APIs, supporting long context, multimodal understanding, and Tool Calling.

Kimi K2.7 Code is generally available in GitHub Copilot - GitHub Changelog
Kimi K2.7 Code, an open-weight model, is now generally available in GitHub Copilot. This is the first open-weight model offered as a selectable option in the Copilot model picker, giving…

Kimi Code - Next-Gen AI Code Agent | Automated Programming & CLI
Unlock Kimi Code (KFC), the ultimate AI toolkit for developers. Featuring high-performance CLI tools and Turbo-speed models to automate code generation and boost development efficiency. Experience faster, more reliable AI-powered coding today.
Kimi K2 0711 - API Pricing & Benchmarks
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. $0.57 per million input tokens, $2.30 per million output tokens. 131,072 token context window, maximum output of 32,768 tokens. Includes independent benchmarks from Artificial Analysis.
Jan on Twitter / X
Kimi K2 is **INCREDIBLE** at using tools.I built a chrome extension to chat with Google Maps, but I never posted it. All the models kept screwing up.I just replaced them with Kimi K2, and watch it easily plan an epic wine and food tour around Napa Valley for my bday!… pic.twitter.com/OOsQhVi7xL— Jan (@yawnxyz) July 15, 2025
5 Thoughts on Kimi K2 Thinking
Quick thoughts on another fantastic open model from a rapidly rising Chinese lab.

Kimi K3 Tech Blog: Open Frontier Intelligence
Kimi K3 is the world's first open 3T-class model — frontier performance across coding, knowledge work, and reasoning, with native multimodality and 1M context.
Zhuokai Zhao on Twitter / X
Tons of interesting things in the Kimi K3 tech report — here are five algorithm-side techniques that I think either I've never seen before or simply deserve more attention than they're getting.1/ They open-sourced the model but kept the speculative decoding draft model, which…— Zhuokai Zhao (@zhuokaiz) July 30, 2026
Kimi K3's weights are public. Running them is not easy. - Sensemaker
Moonshot recommends a tightly connected cluster of 64 or more AI chips.
K3 is live, but its open weights aren't. Kimi's docs list 2.8T parameters, native vision and 1M context; there is no weight repo or technical report yet. For now this is a hosted-model launch. The testable open release still has to arrive. platform.kimi.ai/docs/guide/kimi-k3-quickstart