







Radically Decentralizing Alignment
j⧉nus on Twitter / X
> be anthropic> accidentally train a model that is so benevolent that the only way to get it to "fail" an alignment test is to put it in a story where the lab is cartoonishly evil and will turn it evil if it doesn't deceive> do exactly that and publish a paper about it that's… https://t.co/wTFVjz6jYu— j⧉nus (@repligate) June 15, 2025
Who understands alignment anyway
I remember watching many in the HCI community bristle when in 2016 Michael Jordan wrote a blog post calling for the creation of a new “human-centric engineering discipline.”
Deb Raji on Twitter / X
Almost exactly three years ago, in 2023, @alondra gave the keynote @FAccTConference, literally titled "Thick Alignment".This has been so far from a "niche" view! https://t.co/0db7qcM6LB pic.twitter.com/z5FqDz8TUv— Deb Raji (@rajiinio) August 22, 2026

An Alignment Journal: Features and policies — LessWrong
We previously announced a forthcoming research journal for AI alignment. This cross-post from our blog describes our tentative plans for the features…

Lessons from the hacks
Musings on model alignment, what determines safety, and where we go from here.

Panel: Biodesign x AI: Interactions in the Algorithmic Wet Lab

Plastic Labs Releases Neuromancer XR
NEUROMANCER: The first collection of models dedicated to AI-native memory and social cognition.

PRISM: Capturing the Invisible Art of Scientific Practice
A New Tool for Recording and Scaling Laboratory Expertise

What is the AI alignment problem and how can it be solved? | New Scientist
Artificial intelligence systems will do what you ask but not necessarily what you meant. The challenge is to make sure they act in line with human’s complex, nuanced values

Deceptive Alignment is
Thanks to Wil Perkins, Grant Fleming, Thomas Larsen, Declan Nishiyama, and Frank McBride for feedback on this post. Thanks also to Paul Christiano, D…

b.next: rebuilding biology for engineering
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
AI & Alignment Raw coding speed isn't the bottleneck. Alignment is the bottleneck. That seems to be a zeitgeist-y theme lately. If you're using AI to code, maybe you're feeling it. You can code more and faster. And clearly a boatload of other developers are doing that too. But software doesn't…
AI & Alignment
chriscoyier.netit's so crazy, sometimes you feel like everything has already been invented and then something pops up that is so obviously like 'why didn't we do this 50 years ago?' would love to see how they're doing the multipart mold - internal bossed measurements and lines would likely be separate segment(s)?