







And How They Stack Up Against Qwen3
Georgi Gerganov on Twitter / X
I think the consensus is that Qwen3.5 is a step change so atm I would recommend explore that, given that it covers a range of sizes suitable for all devices.Note that the main issues that people currently unknowingly face with local models mostly revolve around the harness and…— Georgi Gerganov (@ggerganov) March 30, 2026
A 30B Qwen Model Walks Into a Raspberry Pi… and Runs in Real Time
ByteShape's device-optimized release of Qwen3-30B-A3B-Instruct-2507.
Verifying gpt-oss implementations
The OpenAI gpt-oss models are introducing a lot of new concepts to the open-model ecosystem and getting them to perform as expected might ta

qualcomm/nexa-sdk
Run frontier LLMs and VLMs with day-0 model support across GPU, NPU, and CPU, with comprehensive runtime coverage for PC (Python/C++), mobile (Android & iOS), and Linux/IoT (Arm64 & x86 Docker). Supporting OpenAI GPT-OSS, IBM Granite-4, Qwen-3-VL, Gemma-3n, Ministral-3, and more.
Awni Hannun on Twitter / X
According to benchmarks Qwen3.5 4B is as good as GPT 4o.GPT 4o came out ~2 years ago (May 2024).Qwen 3.5 4B runs easily on modern mobile devices.So the gap between frontier intelligence in a datacenter and running a model of equal quality on your iPhone could be 2-3 years.…— Awni Hannun (@awnihannun) March 6, 2026
Eric on Twitter / X
Qwen3.5 27B is awesome (the entire family above 9B is impressive). You can now try it directly in your browser at SOTA speeds with whatever GPU you have: https://t.co/avWxUd8vNLMy previous research in practice - The `Intel/Qwen3.5-27B-int4-AutoRound` is particularly good. https://t.co/QVSVNKM0Nk— Eric (@Ex0byt) March 21, 2026
Artur Chakhvadze on Twitter / X
We are releasing our first quantized checkpoints for the Qwen3.5 series of models, co-designed jointly with our inference engine to achieve maximum possible performance on Apple hardwareStarting from 0.8B, 2B and 4B modelshttps://t.co/2R8BdhAfzv— Artur Chakhvadze (@norpadon) June 8, 2026
clem 🤗 on Twitter / X
The main breakthrough of GPT-5 was to route your messages between a couple of different models to give you the best, cheapest & fastest answer possible.This is cool but imagine if you could do this not only for a couple of models but hundreds of them, big and small, fast and… pic.twitter.com/ww2ApYTN3D— clem 🤗 (@ClementDelangue) October 17, 2025

GPT-5.5: Capabilities and Reactions
The system card for GPT-5.5 mostly told us what we expected.

Building with Open Models
GPT-5.6 Preview System Card - OpenAI Deployment Safety Hub
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch -- our most robust yet -- are built to deliver these models safely and at scale, around the world.

Casper Hansen on Twitter / X
Qwen3.5 Small models about to release!Qwen3.5 9B, 4B, 2B, 0.8B, or something in between is possible.- imagine 9B beating Qwen3-Next-80B- or 4B beating Qwen3-VL-30B in multimodal reasoningBuying a GPU is starting to have high return of intelligence on investment— Casper Hansen (@casper_hansen_) March 1, 2026
Introducing gpt-oss
We’re releasing gpt-oss-120b and gpt-oss-20b—two state-of-the-art open-weight language models that deliver strong real-world performance at low cost. Available under the flexible Apache 2.0 license, these models outperform similarly sized open models on reasoning tasks, demonstrate strong tool use capabilities, and are optimized for efficient deployment on consumer hardware.

GPT-5.6 System Card - OpenAI Deployment Safety Hub
GPT-5.6 is a new family of three models: Sol, our new flagship model; Terra, a capable lower-cost option; and Luna, our fastest and most cost-efficient model. The safeguards we have built for this launch—our most robust yet—are built to deliver these models safely and at scale, around the world.

Previewing GPT-5.6 Sol: a next-generation model
OpenAI previews GPT-5.6 Sol, a next-generation model with stronger capabilities in coding, science, and cybersecurity, paired with its most advanced safety stack.

Qt Platform Abstraction | Platform Integration | Qt 6.11.1
The Qt Platform Abstraction (QPA) is the main platform abstraction layer in Qt.