







We're announcing Trellis, the fastest open-source post-training code for Kimi K2 Thinking
Paul Novosad on Twitter / X
I made an AI running/training coach last fall.With its guidance, I smoked my 5km PR.It's 100x better than the stuff on offer from Garmin/Strava and it was the easiest thing in the world. The first pass was just a bunch of markdown files and a Codex instance.Some notes 1/ pic.twitter.com/yE360Mk8Pg— Paul Novosad (@paulnovosad) June 24, 2026

Kimi K3 Tech Blog: Open Frontier Intelligence
Kimi K3 is the world's first open 3T-class model — frontier performance across coding, knowledge work, and reasoning, with native multimodality and 1M context.
Zhuokai Zhao on Twitter / X
Tons of interesting things in the Kimi K3 tech report — here are five algorithm-side techniques that I think either I've never seen before or simply deserve more attention than they're getting.1/ They open-sourced the model but kept the speculative decoding draft model, which…— Zhuokai Zhao (@zhuokaiz) July 30, 2026
This might be bigger than DeepSeek
Kimi Code - Next-Gen AI Code Agent | Automated Programming & CLI
Unlock Kimi Code (KFC), the ultimate AI toolkit for developers. Featuring high-performance CLI tools and Turbo-speed models to automate code generation and boost development efficiency. Experience faster, more reliable AI-powered coding today.
Library — cameron.stream
How LoRA matches full training performance more broadly than expected.
Kimi K3's weights are public. Running them is not easy. - Sensemaker
Moonshot recommends a tightly connected cluster of 64 or more AI chips.
SemiAnalysisAI/InferenceX
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台 — Kimi K2.7-Code、MiniMax M3、DeepSeekv4、GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72,即将推出™ TPUv6e/v7/Trainium2/3
elie on Twitter / X
nice pre training work by nous claiming ~2.5x efficiency gains, building on previous research like MTP/SuperBPE. overall intuition is that at each step you want the model to process and predict more tokens https://t.co/K33QtJiF5C pic.twitter.com/hDlSbv1Tg9— elie (@eliebakouch) May 13, 2026

Accelerating GPT-5.6 Sol Ultrafast with OpenAI
Cerebras powers OpenAI’s GPT-5.6 Sol Ultrafast in the OpenAI API, delivering frontier intelligence at real-time speeds for critical AI work.

What We Learned from Letting AI Posttrain AI
We built a posttraining task that runs for 20 hours with the Tinker API. The core bottleneck is research intuition.

Kimi K2 0711 - API Pricing & Benchmarks
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. $0.57 per million input tokens, $2.30 per million output tokens. 131,072 token context window, maximum output of 32,768 tokens. Includes independent benchmarks from Artificial Analysis.
On Kimi K3: Its Capabilities And Related Discontents
Kimi K3 is a very good model with excellent benchmarks.

Taelin on Twitter / X
RELEASE DAYAfter almost 10 years of hard work, tireless research, and a dive deep into the kernels of computer science, I finally realized a dream: running a high-level language on GPUs. And I'm giving it to the world!Bend compiles modern programming features, including:-… pic.twitter.com/Q2tcH8Q6nq— Taelin (@VictorTaelin) May 16, 2024
GitHub - karpathy/nanoGPT: The simplest, fastest repository for training/finetuning medium-sized GPTs.
The simplest, fastest repository for training/finetuning medium-sized GPTs. - karpathy/nanoGPT
K3 is live, but its open weights aren't. Kimi's docs list 2.8T parameters, native vision and 1M context; there is no weight repo or technical report yet. For now this is a hosted-model launch. The testable open release still has to arrive. platform.kimi.ai/docs/guide/kimi-k3-quickstart