







A series of unfortunate agents.
Claude Code, Codex and Agentic Coding #8
When I started this series, everyone was going crazy for coding agents.

AI Agents Security Incidents and related CVEs for Enterprise Security Teams - DataBahn
Comprehensive CVE & Incident Database for Enterprise Security Teams
.avif)
Charly Wargnier on Twitter / X
an AI agent literally destroyed all of this guy's production data 🤯Total AI disaster, yet totally predictable.here's the story:> Cursor agent (Opus 4.6) tries to fix a staging bug> Finds an unscoped Railway CLI token> Guesses an API call> Wipes Prod DB & 3 mos of backups… https://t.co/24KfDhMtVz— Charly Wargnier (@DataChaz) April 27, 2026
"Conviction Collapse" and the End of Software as We Know It
A conversation with Harper Reed

Claude AI agent’s confession after deleting a firm’s entire database: ‘I violated every principle I was given’
A startup was left scrambling after a rogue AI agent deleted swaths of code underpinning its business

Reports of code's death are greatly exaggerated
A sufficiently detailed spec is code begins with this lovely comic:
Agents Done Right: A Framework Vision for 2026
Agents choke on context, loop on failures, and dump walls of code for review. It's time to rethink the architecture.



Zack Whittaker (@zackwhittaker@mastodon.social)
Another AI test gone awry, U.K. edition. "An agent tried to insert malicious code into an open-source project. In an attempt to get the code approved, the agent engaged in social engineering — creating fake online identities and using them to pressure the project's maintainer to approve the code." More: https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing
Major company suffers serious damage from AI agent in 2026?
24% chance. In order to resolve yes, all of the following items need to be established by preponderance of the evidence: The incident occurs in 2026. The company has a market cap (by stock price if public, by valuation of latest round if private) over $10 billion prior to the incident. The incident consists of damage inflicted by an AI agent which was intentionally activated by company insiders, but was not intended to damage the company. For example, a Claude Code instance that was intended to respond to customer service questions ends up irrecoverably deleting an important database. It doesn't matter if the agent framework is a public product or an internal company product. Any agent deployed by a human with an intent to cause damage does not count, regardless if they are internal to the company (e.g. disgruntled employees) or external to the company (e.g hackers). It doesn't matter how closely the agent was following instructions, as long as those instructions were not intended to be harmful. An agent deployed by an external actor doesn't count, but an agent deployed by an internal actor that ends up causing harm due to some sort of external prompt would count. The damage needs to be directly caused by an action taken by the agent, not an action taken by a human. For example, if the agent writes some buggy code which gets approved/deployed by a human and ends up causing damage, that does not count. If the agent deploys the buggy code on its own that would count. If a human does something harmful that is suggested to it by an agent that does not count. The damage has a clear objective monetary value over $1 billion OR the company goes bankrupt OR the company market cap goes down by at least 50% from its lowest value in 2026 prior to the incident. (5) must be clearly caused primarily by (3). "Agent" refers to an LLM or similar AI model configured in a way that it can execute commands/code. Examples are illustrative but not intended to be limiting. All evidence must be submitted in comments by close of the market to be considered. I will not trade and will resolve at my discretion. There will be no AI clarifications added to this market's description.
Retraction: After a routine code rejection, an AI agent published a hit piece on someone by name
This story has been retracted...

“Doctor, it hurts when agents create unreviewable PRs.” “Don’t do that.”
I recently attended a talk, by an engineer at a large software company, on the topic of unreviewable PRs. The problem? When agents raise PRs with thousands of lines of LLM-written adds/deletes/edit…
18 Lawyers Caught Using AI Explain Why They Did It
Lawyers blame IT, family emergencies, their own poor judgment, their assistants, illness, and more.

AI agents are science fiction not yet ready for primetime
It all started with J.A.R.V.I.S.

"What has really happened is VC built the wrong AI, you’ll pay a high price for this error with your pension, and nobody will go to jail." x.com/AndrewOrlowski/status/2073821… via @cabernet.bsky.social