







For the past few weeks, I've been playing with gpt-5.6-luna. It is shockingly capable, fast, and smart. I regularly see it do ~100 tps, and rip around my codebase, email, and knowledge base.

Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster. Powered by Cerebras, it delivers up to 750 output tokens per second.

clem 🤗 on Twitter / X
The main breakthrough of GPT-5 was to route your messages between a couple of different models to give you the best, cheapest & fastest answer possible.This is cool but imagine if you could do this not only for a couple of models but hundreds of them, big and small, fast and… pic.twitter.com/ww2ApYTN3D— clem 🤗 (@ClementDelangue) October 17, 2025

OpenAI introduces 'Ultrafast,' a new mode that makes GPT-5.6 Sol work at 14x the speed | TechCrunch
OpenAI is launching a preview of a sped up version of its latest, most powerful model, in an effort to court enterprise users.

Better Call Sol The Workhorse
OpenAI’s GPT-5.6-Sol is finally here, along with the cheaper Terra and Luna.

GPT-5.6 Preview System Card - OpenAI Deployment Safety Hub
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch -- our most robust yet -- are built to deliver these models safely and at scale, around the world.

GPT-5.6 System Card - OpenAI Deployment Safety Hub
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.

GPT-5.5: Capabilities and Reactions
The system card for GPT-5.5 mostly told us what we expected.

Accelerating GPT-5.6 Sol Ultrafast with OpenAI
Cerebras powers OpenAI’s GPT-5.6 Sol Ultrafast in the OpenAI API, delivering frontier intelligence at real-time speeds for critical AI work.

Lee Robinson on Twitter / X
Watch a timelapse of GPT-5.2 building a browser!Both websites start out barely working, and then after millions of lines of code, the browser actually works.Pretty cool experiment with long-running agents. https://t.co/KT6HgHEivA pic.twitter.com/8EPbyih8cU— Lee Robinson (@leerob) January 14, 2026
GPT is the Heroku of AI - Ken Kantzer's Blog
I read a comment on HN that sparked this article: GPT is kind of like DevOps from the early 2000s. Here’s the hot take: I don’t see the primary value of GPT being in its ability to help me develop novel use cases or features – at least not right now. The primary value is […]
Model Release Notes | OpenAI Help Center
We’re beginning the rollout of GPT-5.6 Sol in ChatGPT, our flagship reasoning model for complex work across coding, research, science, cybersecurity, computer use, and design.GPT-5.6 Sol is rolling out to eligible paid ChatGPT plans. Free, Go, and logged-out users are not included. Availability may vary during rollout, and managed-workspace access can depend on administrator settings. Availability for other GPT-5.6 family models varies by product and plan; check the model picker or current rate card in the product you use.

Verifying gpt-oss implementations
The OpenAI gpt-oss models are introducing a lot of new concepts to the open-model ecosystem and getting them to perform as expected might ta

TPUs vs. GPUs and why Google is positioned to win AI race in the long term | Hacker News
To quote The Next Platform: "An Ironwood cluster linked with Google’s absolutely unique optical circuit switch interconnect can bring to bear 9,216 Ironwood TPUs with a combined 1.77 PB of HBM memory... This makes a rackscale Nvidia system based on 144 “Blackwell” GPU chiplets with an aggregate of 20.7 TB of HBM memory look like a joke."