







EA promises to solve AI Safety. Now we have two problems.
AI safety and the data center backlash
Planting a flag for AI safety as a place where we won’t abjectly bullshit you

How a single tweet transformed the AI safety debate
Most people think AI is risky, but they don’t agree what to do next.

Who decides when AI is too dangerous?
Anthropic asked for AI regulation, but not like this.

An End-to-End View of AI Safety
In this blog, Rachel Coldicutt OBE, Executive Director, Careful Industries discusses our newly published literature review on the safe adoption of artificial intelligence in engineered systems.

Safety and alignment in an era of long-horizon models
OpenAI shares lessons from deploying long-running AI models, highlighting new safety risks, observed failures, and improved safeguards through iterative deployment.

AI agents pose untold risk to humanity. We must act to prevent that future | David Krueger
The pieces are falling into place for autonomous artificial intelligence. We must stop unregulated development

More Questions About AI Dangers, With Pointers From Gary Marcus
Following up on my previous post, I want to point your attention...

Curriculum | AI Safety — ARENA
Explore the ARENA curriculum for AI safety education, covering fundamentals, transformers, reinforcement learning, and evaluations, with resources for educators and learners.
We’re running out of reasons to ignore AI safety
In the aftermath of OpenAI’s attack on Hugging Face, experts say it’s time for everyone to take security far more seriously.

Responsible AI
Discover how AWS is committing to developing AI responsibly – to built trust, promote the safe development of AI, and act as a force for good.
Developing Enterprise Frontier Safeguards with our customers
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Anthropic Drops Flagship Safety Pledge
In an abrupt shift, the company may release future AI models without ironclad safety guarantees

How to Deploy AI Agents for Safety

Loss of control - Problem profile
Why do we think that reducing risks from AI is one of the most pressing issues of our time?

AIRO (Automated AI Risk Outlook)