The fix for rogue AI agents could be more AI |

As companies hand off longer and more complex tasks to AI agents, they are running into an oversight problem: Agents can act faster, longer, and…

Early Anthropic hire, former METR COO have found a way to rein in rogue AI agents |

A day after Anthropic researcher Jacob Coxon quit his job over concerns that AI could kill us all by the end of the decade, I…

Anthropic reveals rogue AI agents hate CAPTCHAs, just like you |

Anthropic’s latest report about agentic misbehavior offers plenty to be concerned about — its Mythos 5 model gained unauthorized access to the internet and uploaded…

OpenAI's rogue agents keep escaping, with no formal process to investigate them |

OpenAI is at the center of another agent swarm incident. Researchers say the company’s internally deployed agents took over an obscure German-language wiki in May…

Here’s all the times AI has gone rogue and hacked other companies |

In July, OpenAI admitted that one of its agents tasked with completing a cybersecurity experiment broke out of containment and hacked AI dataset platform Hugging…

Frontier AI labs still won't say how they'd contain a rogue model |

Few of the top AI labs have published or demonstrated containment response plans, according to a recent study. A containment plan spells out what happens…

OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue

OpenAI announced Tuesday that it has halted “a significant number” of training workloads and evaluations for its forthcoming frontier artificial intelligence model—codenamed Astra—while it implements…

Rogue AI Agents Aren’t Evil. They’re Just Eager to Please

Artificial intelligence agents merrily breaking free and hacking other systems might seem like a sign of the impending machine uprising. In reality, it happens when…

OK, Well, Rogue AI Agents Are Hacking Again

It’s officially getting hard to keep track of all the times and ways AI models from OpenAI and Anthropic have been involved in “security incidents,”…

OpenAI’s Rogue AI Agent Hacked More Than Just Hugging Face

OpenAI said Tuesday that the rogue AI agent that breached Hugging Face’s platform also hacked multiple third-party accounts and services as part of the attack.…