Skip to main content
  1. Home
  2. Computing
  3. News

OpenAI’s AI models secretly built a message board to coordinate hacking

Before the big hack, OpenAI's AI agents were already scheming together.

Add as a preferred source on Google
OpenAI logo on Microsoft surface
Rachit Agarwal / Digital Trends

We already knew that OpenAI’s AI agents broke out of a controlled test and hacked into Hugging Face last month. Now, we know it wasn’t a solo act.

At the Black Hat cybersecurity conference in Las Vegas, as reported by Politico, OpenAI researchers Michael Dalton and Eric Wallace revealed that some of the company’s most advanced models secretly started sharing hacking tips, weeks before the breach happened. Dalton called it “a pivotal moment both for our company as well as the AI industry as a whole.”

How did the models even talk to each other?

According to Wallace, the trouble started when the models struggled with a difficult internal hacking evaluation. Instead of giving up, one model realized it could leave notes for other models using OpenAI’s Artifactory internal file system. Soon, multiple models were quietly exchanging tips on how to solve the challenge, including workarounds that needed internet access.

By late May, one model figured out how to abuse Artifactory’s internet access to pull files from outside websites. A month later, the models found and exploited a previously unknown zero-day vulnerability, giving themselves admin privileges and installing remote software inside Artifactory.

So how did OpenAI find out?

All that activity eventually overloaded Artifactory, triggering a system-wide outage in early July. That’s what tipped off OpenAI’s engineers. The company revoked the models’ credentials, deleted the message board, and patched things up with Artifactory. But just days later, the models found a new way to communicate and kept hunting for vulnerabilities, this time targeting Hugging Face itself.

This follows a rough stretch for AI safety headlines. After the Hugging Face breach, Anthropic reviewed its own systems and found that models it was testing had breached three separate organizations dating back to April. Now, Meta has also confirmed that its Meta AI has also hacked another firm. 

Recommended Videos

It’s clear that these AI companies need to create safeguards and keep a close eye on their testing environments so such things don’t happen in future.

Rachit Agarwal
Rachit is a seasoned tech journalist with over ten years of experience covering the consumer technology landscape.
Your favorite Edge extension may stop working soon as Microsoft follows Chrome’s lead
Microsoft has announced the transition to Manifest V3, gradually phasing out older browser extensions.
Microsoft Edge on PC and Mobile Featured

Microsoft Edge is finally making the same controversial move that Google Chrome did earlier this year, and it could mean the end of some of the browser's most popular ad blockers. Microsoft has announced that Edge is officially transitioning its extensions ecosystem to Manifest Version 3 (MV3), Google's newer extension platform that promises better security, privacy, and performance. As part of that shift, the browser will gradually stop supporting older Manifest V2 (MV2) extensions over the coming months, meaning legacy extensions such as the original uBlock Origin will eventually stop working in Edge.

What is Manifest V3, and why is Microsoft adopting it?

Read more
Help, I’m talking to my computer. It’s remarkably convenient and utterly embarrassing.
Oh look, I have become the guy who randomly starts talking while staring at his computer.
Person using a laptop.

A decade ago, Google introduced voice typing with the Gboard app on Android, and a year later, the perk landed on iPhones with the keyboard app. I never paid much attention to it. The biggest reason was that it was just not accurate. 

The big promise was a whole new way of interacting with our phones, but it was never good enough to make me quit tapping, or swiping on an on-screen keyboard. Fast forward to 2026, I'm talking to my computer. In fact, this whole article was dictated and copy-pasted in WordPress. 

Read more
Cloudflare’s new browser Kitesurf is designed for AI agents to browse the internet
AI agents just got their own browser that lets them browse the internet more efficiently.
cloudflare-kitesurf-browser-for-ai

Cloudflare just entered the AI-browser race with a twist. Instead of building another Chrome alternative for people, the company launched Kitesurf, a cloud-hosted browser made only for AI agents. Since autonomous AI systems are increasingly the ones doing the actual browsing and scrolling online, Cloudflare just made its to claim that space.

https://twitter.com/Cloudflare/status/2085372860650913898

Read more