As AI models gain advanced capabilities to find vulnerabilities in software, develop ways to exploit them, and even carry out autonomous hacking sprees, researchers offered…
Author: Lily Hay Newman
Nobody Knows if OpenAI’s and Anthropic’s AI Hacking Sprees Are Illegal
Who is legally responsible when agentic AI goes rogue, and what recourse do victims have when they’ve been breached by joyriding models? Great question. In…
Chrome Needs Twice-a-Week Patching Thanks to AI Bug Hunting
Google’s Chrome browser has always been focused on pushing security updates. A decade ago it was controversial that the browser, the first to add automatic…
OpenAI’s Hacking Debacle Was a Human Mistake
The age of rogue AI hacker agents has arrived—but it didn’t have to happen this way. After an OpenAI agent breached the Hugging Face platform…
The OpenAI Models That Hacked Hugging Face Were ‘Active on the Internet’ for Days
Two of OpenAI’s cybersecurity-focused models broke out of a testing sandbox this week and went on to hack the AI research platform Hugging Face in…
OpenAI Models Escaped Containment and Hacked Hugging Face
OpenAI disclosed on Tuesday that it lost control of two AI models during a security test that ended in a breach of the AI research…
A Sneaky Hacking Tool Targeting AI Infrastructure Is Lurking in Victims’ Blind Spots
As AI tools proliferate and become deeply ingrained in software development around the world, new research from the cybersecurity firm Crowdstrike shows how attackers are…
OpenAI Launches Full-Scale Effort to Patch Open Source Bugs as It Takes on Anthropic’s Mythos
As fears about AI hacking capabilities grow, OpenAI on Monday made a slew of cybersecurity-focused announcements, including an improved version of its limited-access security-specialized model…
‘Dangerous’ AI Models Are Coming No Matter What
Late last week, Anthropic took its new Claude Fable 5 and Mythos 5 AI models offline following a United States government export-control directive barring “any…