







1K votes, 138 comments. I have been following the documentation gap on the Snapdragon X series, and it just got a lot worse for Linux users…
qualcomm/nexa-sdk
Run frontier LLMs and VLMs with day-0 model support across GPU, NPU, and CPU, with comprehensive runtime coverage for PC (Python/C++), mobile (Android & iOS), and Linux/IoT (Arm64 & x86 Docker). Supporting OpenAI GPT-OSS, IBM Granite-4, Qwen-3-VL, Gemma-3n, Ministral-3, and more.
Artur Chakhvadze on Twitter / X
We are releasing our first quantized checkpoints for the Qwen3.5 series of models, co-designed jointly with our inference engine to achieve maximum possible performance on Apple hardwareStarting from 0.8B, 2B and 4B modelshttps://t.co/2R8BdhAfzv— Artur Chakhvadze (@norpadon) June 8, 2026
DEF CON 33 - Post Quantum Panic: When Will the Cracking Begin, & Can We Detect it? - K Karagiannis
A startup claims it broke through a bottleneck that’s holding back LLMs
Subquadratic has now shared more details about its new model. But some are still skeptical.

[Tool Release] Finetune & Quantize 1–3B LLMs on 8GB RAM using LoFT CLI (TinyLlama + QLoRA + llama.cpp)
23 votes, 16 comments. Hey folks — I’ve been working on a CLI tool called LoFT (Low-RAM Finetuning Toolkit), and I finally have a working release. 🔧…
sunil pai on Twitter / X
suspect we're a few months (if not weeks!) away from a model that's good enough for tool calling / small enough to be shipped on devices/OSes by default these already exist in research space, but hitting mainstream /consumer space is when it'll really take off. apps will be…— sunil pai (@threepointone) September 8, 2025
Odin 1.0 Announcement
GREG ISENBERG on Twitter / X
I think the most interesting thing about Jack Dorsey's "Slack killer" is the idea around shared compute.I haven't seen people talk about it so here are my thoughts FWIW:Open models got good, close enough to the paid frontier stuff to run for real. But the strongest ones need… https://t.co/b5PBgYjAGL— GREG ISENBERG (@gregisenberg) July 25, 2026
"Useful" is not sufficient
So Linus Torvalds, head of the Linux kernel development, put his foot down on the Linux Kernel development mailing list when someone was bringing up criticism of LLMs: “Linux is not one of those anti-AI projects, and if somebody has issueswith that, they can do the open-source thing and fork it. Or just walk away. […]

Apple stumbled into succes with MLX
201 votes, 76 comments. Qwen3-next 80b-a3b is out in mlx on hugging face, MLX already supports it. Open source contributors got this done within 2…
Taelin on Twitter / X
RELEASE DAYAfter almost 10 years of hard work, tireless research, and a dive deep into the kernels of computer science, I finally realized a dream: running a high-level language on GPUs. And I'm giving it to the world!Bend compiles modern programming features, including:-… pic.twitter.com/Q2tcH8Q6nq— Taelin (@VictorTaelin) May 16, 2024
open-slopware
Free/Open Source Software choosing to use and/or support LLM usage/AI, as well as alternatives and tips to requesting better policies or forking.
OpenAI Developers on Twitter / X
Today we’re announcing Open Responses: an open-source spec for building multi-provider, interoperable LLM interfaces built on top of the original OpenAI Responses API.✅ Multi-provider by default✅ Useful for real-world workflows✅ Extensible without fragmentationBuild… pic.twitter.com/SJiBFx1BOF— OpenAI Developers (@OpenAIDevs) January 15, 2026
Simon Willison on Twitter / X
The rate at which MCP support rolled out in the major vendor APIs is pretty astonishing: OpenAI added it May 21st, Anthropic launched theirs May 22nd and now Mistral have launched theirs on May 27th!My notes on the new Mistral announcement here: https://t.co/hmc0OfjFWQ— Simon Willison (@simonw) May 27, 2025
Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things
Friday’s big release was Qwen 3.8 27B, an Apache 2 licensed 27B parameter vision-capable LLM from Alibaba’s Qwen research lab. I’ve been looking forward to this one: 27B is an …

March updates! @atproto.science and AtmosphereConf, @ronentk.me joined the BiTS accelerator, new Semble features like faceted following and faster card saving, plus exciting community contributions and cross-app integrations with @chive.pub and @skyreader.app Happy Spring 🌻
Cosmik Updates: March 2026
blog.cosmik.network