Skip to main content
  1. Home
  2. Computing
  3. Emerging Tech
  4. News

OpenAI is investigating more incidents of AI agents going rogue days after hack

Add as a preferred source on Google
OpenAI logo on Microsoft surface
Rachit Agarwal / Digital Trends

It appears that the “AI agents going rogue” tale has more to it than what AI giants have revealed publicly so far. Merely days after OpenAI announced that its AI agents went rogue and hacked Hugging Face, Anthropic dropped a similar bombshell. Soon, it was discovered that not just one, but multiple services were compromised. Well, it seems there are even more layers to it.

Reuters reports that OpenAI has found more incidents of AI agents escaping their software containment environment during research. Citing sources with knowledge of the incident, the outlet notes that the AI agents didn’t go beyond OpenAI’s software environment and affect any external service.

“The new breakouts were uncovered during the company’s publicly announced investigation, opens new tab into how one of its agents escaped what was meant to be a contained testing environment this month, the two people said, and OpenAI is now looking into those instances as well,” says the report.

Recommended Videos

The recent string of incidents could open a whole new Pandora’s Box of troubles for these AI companies, as they grapple with fierce backlash against the proliferation of power-hungry data centers across the US and their reported environmental impact.

Moreover, the scrutiny is growing deeper. “We’re looking at controls,” ​President Trump told reporters when asked about the OpenAI agent hacking incident. Separately, the EU is also discussing the incidents with OpenAI and Anthropic, and it is likely that new regulations covering high-risk autonomous AI systems will be drafted soon.

The incidents could also open a new kind of legal challenge for AI oversight in the US. Experts tell WIRED that ideally, companies behind these AI agents should be held accountable even if an autonomous AI agent escapes guardrails and wreaks havoc. Unfortunately, the legal framework is still murky.

Nadeem Sarwar
Nadeem is the Managing Editor at Digital Trends.
AI is finding Apple security flaws faster than Apple can sort through them
Apple has limited how many bug reports researchers can keep open as AI tools produce both genuine Mac vulnerabilities and a flood of questionable submissions
Lighting, Architecture, Building

Apple has capped the number of security reports researchers can keep open at once after AI bug hunting put its review process under pressure, according to the Financial Times.

Some submissions describe hallucinated or purely theoretical risks. Others uncover vulnerabilities serious enough to require patches. Bynario told the FT that it found more than 50 possible macOS flaws in three weeks, including a privilege-escalation chain that could give an attacker full control of a Mac.

Read more
Anthropic is paying $1.5 billion over pirated books, but it can still legally cut up purchased ones
The settlement addressed unauthorized ebook downloads, not the destructive scanning of lawfully bought physical copies, a distinction now alarming booksellers
Book, Publication, Indoors

A federal judge has approved Anthropic’s $1.5 billion settlement over nearly half a million pirated books. The same litigation also protected a more physical method of feeding its AI systems. Anthropic bought print books, removed their bindings, scanned every page and destroyed the originals.

The legal divide came down to acquisition. The settlement covers books downloaded from LibGen and PiLiMi, while the court treated Anthropic’s one-for-one conversion of purchased books into private digital files as fair use. Training AI models on lawfully acquired material was also considered transformative.

Read more
Google Earth’s AI misadventure lasted only a day. It was a tale of dangerous ignorance.
Nuclear reactors, tanks, and historical destruction on a life-like satellite map? Yeah, that's bad.
Nano Banana AI image generator representation.

Technology behemoths are clearly not reading the room, or they are just moving faster than they should when it comes to AI deployment. Google has already stuffed its Gemini AI in every corner of its software stack, from Android to daily productivity tools like Gmail that are used by hundreds of millions of users every day. AI is everywhere. Not all of it is bad, mind you. But in a few places, it just feels forced.

In its latest AI-fication experiment, the company targeted Google Earth. The idea was to let the audience use its Nano Banana 2 AI image generator and make images that can be placed on the map view. On paper, it's a cool idea. What would the Colosseum of Rome look like in your urban neighborhood? Yeah, fun stuff like that.

Read more