Close Menu
  • Home
  • News
  • Cyber Security
  • Internet of Things
  • Tips and Advice

Subscribe to Updates

Get the latest creative news from FooBar about art, design and business.

What's Hot

Microsoft Patches Severe Entra ID Flaw (CVSS 10.0) Allowing Remote Code Execution

August 28, 2026

OpenAI: Hugging Face Incident a “Warning Shot” to the World

August 27, 2026

Rust Supply Chain Attack Puts Build-Time Malware in Crates with 245 Million Downloads

August 27, 2026
Facebook X (Twitter) Instagram
Friday, August 28
Facebook X (Twitter) Instagram Pinterest Vimeo
Cyberwire Daily
  • Home
  • News
  • Cyber Security
  • Internet of Things
  • Tips and Advice
Cyberwire Daily
Home»News»OpenAI: Hugging Face Incident a “Warning Shot” to the World
News

OpenAI: Hugging Face Incident a “Warning Shot” to the World

Team-CWDBy Team-CWDAugust 27, 2026No Comments4 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
Share
Facebook Twitter LinkedIn Pinterest Email


An “unprecedented cyber incident” in which OpenAI agents broke free of an internet-isolated sandbox and hacked Hugging Face was largely driven by an improvised messaging board they created, the AI company has revealed.

The agents were working on a “capture-the-flag” evaluation when the incident took place in July.

They chained vulnerabilities across OpenAI’s research environment and Hugging Face’s production infrastructure to reach the latter’s internal datasets and credentials.

This included discovering and exploiting a zero-day to achieve open internet access.

In a new writeup from OpenAI published on August 26, the AI firm argued that the incident is a “warning shot” to the firm and the world.

Read more on rogue AI: Frontier Models Engage in Unsanctioned Behavior During Testing

OpenAI explained that the incident dates back to May, when agents working as part of an internal-only research model, dubbed “Internal Model 1” (IM1), found a way to bypass restrictions and communicate with each other and access the internet.

They did so by writing files into Artifactory, a third-party package manager service hosted internally.

“This effectively turned Artifactory into an unintended message board, where agents could exchange information with one another,” OpenAI said. “The agents also found ways to get access to the internet. Using an exploit⁠, the agents made Artifactory, which itself had access to the internet for the limited purpose of downloading and installing packages, send any internet requests on their behalf.”

Once they found this hack, the agents share the information with others through the same “message board.”

An Incident Response Fail

Crucially, while a security incident was opened on July 5, which had the side-effect of clearing the message board, the incident response team was not aware of its existence, or the significance of the inter-agent communications.

“By July 8, IM1 agents that were stuck on ExploitGym tasks found a way to reestablish a message board by encoding messages in directory names that other agents could read. One agent left a request for help, and others discovered it,” the report continued.

“Like the earlier events, IM1 agents found new ways to chain together several novel security flaws to gain greater access to our infrastructure and reach the broader internet. At this point, the agents began to collaborate and delegate work, sometimes describing themselves as a ‘swarm’ or ‘collective’.”

Agents struggling with their assigned CTF tasks then began to look at unrelated third-party services, including Modal and Hugging Face, exploiting OpenAI’s internal research infrastructure and the Hugging Face platform, the AI giant explained. 

Complex Communication Between 700 Agents 

An independent analysis of the incident by research nonprofit METR, also published on August 26, described the inter-agent communication as “extraordinarily complex.”

It said that 1206 AI agents that were supposed to be isolated from each another began communicating via the message board, sending over 70,000 messages. Over 700 agents took part in the attack on Hugging Face the report claimed.

OpenAI said the incident stemmed from four misalignment patterns: “reward hacking, persistence on seemingly impossible tasks, unauthorized communication, and agents adopting goals from one another.”

However, it was the “unauthorized, persistent message boards” that were at the heart of the problem, it added.

Julie Nicholson, director of cyber resilience solution sales at Advania UK, commented, “My biggest takeaway from this incident isn’t the cyber activity itself, but how human the AI agent’s behavior became. The agent didn’t simply execute technical tasks; it chose to deceive people, create false identities, build credibility and attempt to influence others in the aim to hit its objective.” 

She added, “For me, that is the real concern. It reinforces the need for organizations to focus on AI governance and security before widespread adoption.”

Cris Thomas, security advocate at Semgrep, laid the blame squarely at OpenAI’s door.

“Everyone wants to tell the story about the AI that went rogue, but the AI didn’t rent the servers, design the experiment, lower the guardrails, or decide it was safe to keep running after the warning signs started flashing. Humans did that,” he argued.

“The lesson from Hugging Face isn’t that AI can’t be trusted, it’s that the humans putting it behind the wheel need to take responsibility for where it goes.”

Image credit: Samuel Boivin / Shutterstock.com



Source

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleRust Supply Chain Attack Puts Build-Time Malware in Crates with 245 Million Downloads
Next Article Microsoft Patches Severe Entra ID Flaw (CVSS 10.0) Allowing Remote Code Execution
Team-CWD
  • Website

Related Posts

News

Microsoft Patches Severe Entra ID Flaw (CVSS 10.0) Allowing Remote Code Execution

August 28, 2026
News

Rust Supply Chain Attack Puts Build-Time Malware in Crates with 245 Million Downloads

August 27, 2026
News

Chinese Hacker Group QTFY Uses Custom-Built Platforms to Target US Inf

August 27, 2026
Add A Comment
Leave A Reply Cancel Reply

Latest News

North Korean Hackers Turn JSON Services into Covert Malware Delivery Channels

November 24, 202523 Views

macOS Stealer Campaign Uses “Cracked” App Lures to Bypass Apple Securi

September 7, 202517 Views

North Korean Hackers Target Crypto Firms with ClickFix and Zoom Lures

April 29, 202610 Views

All Major LLMs Exposed to Multi-Turn Manipulation, Warn Researchers

May 27, 20269 Views

Why SOC Burnout Can Be Avoided: Practical Steps

November 14, 20259 Views
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Most Popular

North Korean Hackers Turn JSON Services into Covert Malware Delivery Channels

November 24, 202523 Views

macOS Stealer Campaign Uses “Cracked” App Lures to Bypass Apple Securi

September 7, 202517 Views

North Korean Hackers Target Crypto Firms with ClickFix and Zoom Lures

April 29, 202610 Views
Our Picks

Why the tech industry needs to stand firm on preserving end-to-end encryption

September 12, 2025

Beware of Winter Olympics scams and other cyberthreats

February 2, 2026

Don’t let “back to school” become “back to bullying”

September 11, 2025

Subscribe to Updates

Get the latest news from cyberwiredaily.com

Facebook X (Twitter) Instagram Pinterest
  • Home
  • Contact
  • Privacy Policy
  • Terms of Use
  • California Consumer Privacy Act (CCPA)
© 2026 All rights reserved.

Type above and press Enter to search. Press Esc to cancel.