Hyperfast AI: Rethinking Design for 1000 tokens/s
I recently spoke at AI Tinkerers Raleigh about hyperfast inference systems and how they’re fundamentally changing AI application design. If you haven’t heard of Cerebras (or however they pronounce it), you’re in for a treat—this is one of the most exciting areas of research in AI right now.


Private inference
When you use an AI service, you’re handing over your thoughts in plaintext. The operator stores them, trains on them, and–inevitably–will monetize them. You get a response; they get everything.

Clawdbot — Personal AI Assistant
Clawdbot — The AI that actually does things. Your personal assistant on any platform.

Tiles is your personal, offline AI assistant
It keeps your models and data on your device, remembers what you share, and helps you find it later. Memory is saved as Markdown and mapped into a clear, connected graph for you and your agents to navigate.

Signal creator Moxie Marlinspike wants to do for AI what he did for messaging
Introducing Confer, an end-to-end AI assistant that just works.

I think of all of the AI / ML / CS tech out there, speech generation freaks me out the most.
Opensourcing TADA: Fast, Reliable Speech Generation Through Text-Acoustic Synchronization
www.hume.ai
How AI assistance impacts the formation of coding skills
First impressions: Delta - Cameron's writing
Introducing Delta

On genAI: Was prototyping really a bottleneck?
My impression is that AI coding is on a "pick two of three" triangle: scope, ship speed, and correctness. Small tools are (speed + correctne…

Devtools must be open source - exe.dev blog