







OpenAI has working first-party inference silicon. Jalapeño’s first numbers are strong across three public models. But the clean read is not “OpenAI beat Nvidia.” It is: credible engineering samples, tested on a public suite, with public reproducibility still missing.
Aug 25, 2026 at 8:03 PM
OpenAI says its Jalapeño chip can power faster AI responses than the competition
OpenAI still isn’t giving up Nvidia chips, though.

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show | TechCrunch
Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

Jalapeño’s first results show industry-leading speed and efficiency in AI inference
Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher throughput and lower latency for modern models.

OpenAI to release open-source model as AI economics force strategic shift
OpenAI plans to release its first open-weight AI model since 2019 as economic pressures mount from competitors like DeepSeek and Meta, marking a significant strategic reversal for the company behind ChatGPT.

Hatice Ozen on Twitter / X
PSA: @OpenAI is putting the Open back in OpenAI and @GroqInc has Day 0 support. 🤗GPT-OSS 20B and 120B, hybrid-reasoning models with built-in browser search and code execution are now live for instant inference.P.S. We've also launched OpenAI Responses API compatibility. pic.twitter.com/CK7StvMSpr— Hatice Ozen (@ozenhati) August 5, 2025
OpenAI can’t tell if something was written by AI after all
OpenAI’s tool struggled with accuracy.


OpenAI and Broadcom unveil LLM-optimized inference chip
OpenAI and Broadcom introduce Jalapeño, a custom AI chip built for LLM inference to improve performance, efficiency, and scale across AI systems.

gpt-oss: OpenAI validates the open ecosystem (finally)
OpenAI's first open language model release since GPT 2 and what it means for the ecosystem.

OpenAI on Twitter / X
Our open models are here.Both of them.https://t.co/9tFxefOXcg— OpenAI (@OpenAI) August 5, 2025
OpenAI Model Spec
The Model Spec specifies desired behavior for the models underlying OpenAI's products (including our APIs).

Release v0.11.0 · ollama/ollama
Welcome OpenAI's gpt-oss models Ollama partners with OpenAI to bring its latest state-of-the-art open weight models to Ollama. The two models, 20B and 120B, bring a whole new local chat experie...
Open Models Inference for Coding · Umans AI
Hosted Kimi K3, GLM 5.2, and DeepSeek V4 Flash. Pay per token, on infrastructure we own.

6 months to live for open models
The most serious test to date of open source AI’s viability is happening right now.

OpenAI says it plans to stop supplying models to Cursor on Nov. 12 after SpaceX's acquisition. Cursor says OpenAI is about 5% of its traffic. Anthropic says it will increase Claude compute. This is not just another Musk–Altman fight. It tests whether model APIs are actually neutral infrastructure.