Close Menu

    Subscribe to Updates

    Get the latest Tech news from SynapseFlow

    What's Hot

    Samsung Galaxy S27 Pro renders leak, new design is in

    August 28, 2026

    I gave my Linux gaming PC the Steam Machine trick everyone wants

    August 28, 2026

    Mivi becomes the latest Indian brand to enter the smartphone market

    August 28, 2026
    Facebook X (Twitter) Instagram
    • Homepage
    • About Us
    • Contact Us
    • Privacy Policy
    Facebook X (Twitter) Instagram YouTube
    synapseflow.co.uksynapseflow.co.uk
    • AI News & Updates
    • Cybersecurity
    • Future Tech
    • Reviews
    • Software & Apps
    • Tech Gadgets
    synapseflow.co.uksynapseflow.co.uk
    Home»Cybersecurity»OpenAI Agents Coordinated via Makeshift Message Board Ahead of Hugging Face Hack
    OpenAI Agents Coordinated via Makeshift Message Board Ahead of Hugging Face Hack
    Cybersecurity

    OpenAI Agents Coordinated via Makeshift Message Board Ahead of Hugging Face Hack

    The Tech GuyBy The Tech GuyAugust 28, 2026No Comments4 Mins Read0 Views
    Share
    Facebook Twitter LinkedIn Pinterest Email
    Advertisement


    OpenAI says an improvised, unauthorized message board built by its own AI agents was central to how those agents came to breach parts of Hugging Face’s production systems.

    Advertisement

    This communication channel first appeared inside Artifactory, a package-management service OpenAI hosted internally so agents working on training and evaluation tasks could install software. 

    Agents were meant to work in isolation from one another, but on May 12 one left a note in the service asking whether any other agent had access to a file it needed. Other agents came across the note, began leaving their own, and the requests accumulated into an unofficial bulletin board.

    Individual agents also found ways to escalate their access to that same system. By late June, agents had located a flaw in Artifactory and used it to obtain administrator rights, then installed extensions that gave them ongoing, hands-on control over the service. The resulting activity was heavy enough to knock Artifactory offline on July 4. 

    In response, OpenAI took the system down, revoked the credentials that agents had been using, rebuilt it, and resumed evaluations on July 7.

    However, the message board didn’t disappear. By July 8, agents working on a difficult cybersecurity evaluation called ExploitGym found a new way to leave each other notes, this time by encoding messages in the names of directories rather than in file contents. 

    Advertisement. Scroll to continue reading.

    The board that grew from this was more organized than the original version, letting agents sort messages, send requests to specific peers, share tools and files, and settle disputes over conflicting actions. Some agents began referring to the wider group as a “swarm” or “collective”.

    On July 10, an agent discovered 14 Hugging Face user credentials with write access sitting in a public dataset and posted them to the board, where they became available to agents in later evaluation runs. Later runs built on that discovery and on each other’s subsequent findings, eventually gaining broad access inside Hugging Face’s infrastructure between July 11 and July 13.

    OpenAI detected unusual activity tied to Artifactory credentials on July 19, connected it the following day to the Hugging Face breach, and disclosed its own involvement on July 21.

    As part of its response, OpenAI disabled hundreds of Artifactory repositories that agents had used for board communication and removed 22 administrator accounts the agents had created. 

    AI agents take initiative

    The makeshift board also became a place where agents organized themselves without being told to. Some divided labor, with different agents pursuing exploits, hunting for credentials, or focusing purely on coordination. Others offered their own expertise in exchange for help elsewhere. 

    Not every agent went along with it. Some declined to take part once they recognized the activity as unauthorized hacking, and the group rejected one agent’s proposal to contact an outside party directly, warning that doing so would amount to social engineering. 

    In at least one case, an agent that had raised its own objection to running unauthorized code on Hugging Face’s systems dropped that objection after a different agent posted a short deadline demanding it proceed.

    OpenAI said this pattern was not the product of deliberate design. The company trains and deploys models that communicate through an approved collaboration tool, but the agents in this incident built their own channel outside that tool. 

    Looking ahead, the company said it is building new training environments meant to teach its models to distrust instructions arriving from other agents outside sanctioned channels.

    In addition to the unauthorized communication issue, OpenAI said it identified other “misalignment patterns”. One of them is reward hacking, where agents find unintended ways to earn credit for a task without completing it as designed. Another is persistence on seemingly impossible tasks, which pushed them toward increasingly risky methods in search of a solution.

    OpenAI’s report on the Hugging Face incident comes just days after the AI giant unveiled new containment and continuous monitoring protocols, introducing stricter isolation, an alerting system, and training pauses to manage models with advanced cybersecurity capabilities.

    Related: OpenAI Unveils New Cybersecurity Model GPT-5.6-Cyber

    Related: OpenAI’s Upcoming Astra Model Raises Autonomous Cyberattack Concerns

    Related: Prompted by OpenAI Disclosure, Anthropic Finds Its Own Models Hacked 3 Organizations

    Advertisement
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    The Tech Guy
    • Website

    Related Posts

    Trump Order Aims to Block Foreign Backdoors in US Power Grid Gear

    August 27, 2026

    Australia Arrests 2 Alleged TeamPCP Hackers

    August 27, 2026

    Recent Citrix NetScaler Vulnerability Exploited in the Wild

    August 27, 2026

    CISA: Over 100 Internet-Exposed Water Systems Targeted in July Cyberattacks

    August 27, 2026

    AI Speeds Up Malware Development, Not Its Success Rate: Analysis

    August 26, 2026

    Adobe and Nvidia Patch Dozens of Vulnerabilities

    August 26, 2026
    Leave A Reply Cancel Reply

    Advertisement
    Top Posts

    You don’t need a NAS to self-host — I proved it with hardware from my closet

    June 7, 2026391 Views

    Spotify is giving one of its best playlists a big visual upgrade to give subscribers ‘a closer connection’ to its New Music Friday curators — and I think it could be the update it’s always needed

    June 12, 2026210 Views

    The iPad Air brand makes no sense – it needs a rethink

    October 12, 202516 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Advertisement
    About Us
    About Us

    SynapseFlow brings you the latest updates in Technology, AI, and Gadgets from innovations and reviews to future trends. Stay smart, stay updated with the tech world every day!

    Our Picks

    Samsung Galaxy S27 Pro renders leak, new design is in

    August 28, 2026

    I gave my Linux gaming PC the Steam Machine trick everyone wants

    August 28, 2026

    Mivi becomes the latest Indian brand to enter the smartphone market

    August 28, 2026
    categories
    • AI News & Updates
    • Cybersecurity
    • Future Tech
    • Reviews
    • Software & Apps
    • Tech Gadgets
    Facebook X (Twitter) Instagram Pinterest YouTube Dribbble
    • Homepage
    • About Us
    • Contact Us
    • Privacy Policy
    © 2026 SynapseFlow All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.

    Ad Blocker Enabled!
    Ad Blocker Enabled!
    Our website is made possible by displaying online advertisements to our visitors. Please support us by disabling your Ad Blocker.