







Anthropic and Alignment
Anthropic is in a standoff with the Department of War; while the company’s concerns are legitimate, it position is intolerable and misaligned with reality.


Doll (@dollspace.gay)
It seems anthropic is catching up to doll six months ago on drift containment being paramount. https://www.anthropic.com/research/assistant-axis
Anthropic has caught up to OpenAI in image understanding
But neither one is all that good.

The shock of the anthropocene: the earth, history and us
Dissecting the new theoretical buzzword of the "Anthrop…

Anthropic on Twitter / X
New Anthropic research: A global workspace in language models.Of everything happening in your brain right now, only a tiny fraction is consciously accessible—thoughts you can describe, hold in mind, and reason with.We found a strikingly similar divide inside Claude. pic.twitter.com/aLUPBifxth— Anthropic (@AnthropicAI) July 6, 2026
j⧉nus on Twitter / X
> be anthropic> accidentally train a model that is so benevolent that the only way to get it to "fail" an alignment test is to put it in a story where the lab is cartoonishly evil and will turn it evil if it doesn't deceive> do exactly that and publish a paper about it that's… https://t.co/wTFVjz6jYu— j⧉nus (@repligate) June 15, 2025
Home \ Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

WSJ Article Claiming China Has Matched Anthropic Is Obvious Nonsense
The Wall Street Journal printed an outright false headline and heavily misleading story claiming this, which of course was uncritically amplified by the usual suspects.

An update on recent Claude Code quality reports
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

Claude’s Constitution
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

Claude’s Constitution
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

No need to panic about Anthropic’s new blog, and some more good news
The twitterverse is all verklempt with Anthropic’s latest blog.

Anthropic expands partnership with Google and Broadcom for multiple gigawatts of next-generation compute
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Who understands alignment anyway
I remember watching many in the HCI community bristle when in 2016 Michael Jordan wrote a blog post calling for the creation of a new “human-centric engineering discipline.”
"This paper prompted Jack Clark, one of the co-founders of Anthropic to post this to their news letter: 'the singularity could be delayed'".
Mark Riedl
Fascinating experiment: current AI systems lack creativity to reliably pursue research arxiv.org/abs/2607.27191 - poor judgment about the bar for publishable research - uncreative responses in research design - ineffective backtracking from dead ends - poor resource awareness - instruction drift