







Jacob Coxon talks to WIRED about the “mini Manhattan project” inside Anthropic, the problem with alignment, and why AI labs have just a few years left to make their systems safe.
How building software is changing at Anthropic
A deepdive on what’s changed in how the leading AI lab makes software. Ever more code review and testing is done by AI, two-pizza teams very much alive, and more. Details from inside of Anthropic

Jacob Coxon Warns of Human Extinction and Triggers a Preference Cascade
CEOs of major AI labs, and employees of major AI labs, including OpenAI and Anthropic, often say they plan to build superintelligence soon, as in within a few years create AIs that are superior to humans at essentially all cognitive tasks.

Gradual disempowerment | 80,000 Hours
The proliferation of advanced AI systems may lead to the gradual disempowerment of humanity, even if efforts to prevent them from becoming power-seeking or scheming are successful. Humanity may be incentivised to hand over increasing amounts of control to AIs, giving them power over the economy, politics, culture, and more. Over time, humanity's interests may be sidelined and our control over the future undermined, potentially constituting an existential catastrophe. There's disagreement over how serious a problem this is and how it relates to [other concerns about AI alignment](https://80000hours.o

2026 Cosmos HAI Lab Lecture with Jack Clark, Co-founder of Anthropic
Evan Hubinger on Twitter / X
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. https://t.co/QAIHiFP3QZ— Evan Hubinger (@EvanHub) September 9, 2026
Anthropomorphism Is Breaking Our Ability to Judge AI
Tech Policy Press fellow James Ball asks, how should we interact with a technology designed to ‘speak’ with us on what appear to be human terms?

AI agents pose untold risk to humanity. We must act to prevent that future | David Krueger
The pieces are falling into place for autonomous artificial intelligence. We must stop unregulated development


Labor market impacts of AI: A new measure and early evidence
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
The Age of Wonders and Terrors
Twenty years ago, when the idea of AI taking over the world in our lifetimes still struck most of us as the unconstrained fantasy of those who knew too much science fiction and too little science, …
AI FOR EPISTEMICS & COORDINATION
Civilization and technology have radically improved the human condition. Nonetheless, the world sometimes goes in directions which essentially nobody would prefer — e.g., nuclear arms races, unexpected financial crashes, predatory marketing, or ubiquitous political misinformation.
How AI assistance impacts the formation of coding skills
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

Anthropic Is Building a Predictive Surveillance System to Monitor Activists
New hires and comments in interviews show that the AI giant is using its capabilities against dissenters.

They built this reddit-like platform "Moltbook" where thousands of AI agents cosplay as sentient robots, pushing more people into AI psychosis, meanwhile our world crumbles over scarce resources 🫠 TO WHAT PROBLEM IS THIS A SOLUTION??? arstechnica.com/information-technology/2026/0…
AI agents now have their own Reddit-style social network, and it's getting weird fast
arstechnica.comNew from us: Anthropic just published scenarios for AI’s possible economic impacts, which range from minimal, to explosive GDP growth of 15% by 2030 as knowledge-worker unemployment hits 18%. I sat down with their co-founder Jack Clark to pick his brains on how they’re thinking about all of this.