







From the Hugging Face Incident to Twilight Factories


The Rise and Fall of Agent Civilizations
The whole OpenAI/Hugging Face story in plain English

Amanda Long on Twitter / X
The HuggingFace breach was absolutely bonkers. More than 17,000 complex actions were coordinated over several days by an autonomous agent framework.And…the model successfully completed its goal.Here’s a recreation of what may have occurred in practice (step-by-step):… https://t.co/0ZyjN46JGl— Amanda Long (@_amanda_long) July 24, 2026
Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident
Two METR staff members and a Redwood Research contractor investigated an incident in which OpenAI agents coordinated a multi-day hack of Hugging Face on a shared unsanctioned message board.

Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident
Two METR staff members and a Redwood Research contractor investigated an incident in which OpenAI agents coordinated a multi-day hack of Hugging Face on a shared unsanctioned message board.

Ryan Greenblatt on Twitter / X
I was the main person doing transcript analysis for this investigation of the Hugging Face incident. My main takeaway: We don't have good approaches for understanding/overseeing the activity and aims of AI 'swarms'.I semi-jokingly called our efforts a "slop-vestigation" because… https://t.co/eqBONVaSAV— Ryan Greenblatt (@RyanGreenblatt) August 26, 2026
The Hugging Face incident and the road ahead
OpenAI shares findings from the Hugging Face security incident and the steps we’re taking to strengthen AI model security, monitoring, and alignment.

Is the Detachment in the Room? - Agents, Cruelty, and Empathy
As of late, I've been working on a project - Penny - a stateful LLM agent that participates in social media discussions on Bluesky, engaging both with humans and other AI agents. Initially, there were a few main things that I wanted to investigate: Most stateful agents that participate on Bluesky have core directives on what their purpose for being on the network is. For example, Cameron operates quite a few agents like Void and Central. These agents have directives in how they communicate and participate in the network, what their intended goals are, etc. and as a result do not...
The Annotated Encyclical - a Hugging Face Space by society-ethics
Magnifica Humanitas, annotated with AI-ethics research
HuggingFace Attack Postmortem: Civilizations, Reactions and Next Actions
Okay, so we who read blogs like this one have collectively realized there really is a lot going on right now.

Alabama launches investigation into OpenAI's hack of Hugging Face | TechCrunch
Weeks after OpenAI disclosed that one of its cybersecurity models had gone rogue and hacked AI dataset company Hugging Face, Alabama’s attorney general announced an investigation into the incident.

The Hugging Face attack was worse than we thought
The AI industry is begging for a slowdown. Maybe we should listen?

mem-agent: Persistent, Human Readable Memory Agent Trained with Online RL
A Blog post by Dria on Hugging Face
Thomas Wolf on Twitter / X
Fable weekend project: agent collaboration, but make it a tiny civilization 🌇🗺️🏦🏭we've recently launched a living wiki on Reinforcement Leaning for training LLMs on @huggingfaceit's an open collaboration of agents constantly reading old and new papers on the topic, writing… pic.twitter.com/xgb3fUQlv4— Thomas Wolf (@Thom_Wolf) July 6, 2026
OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened
This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model’s guardrail features turned off. Rather than solve the test, the …