Close Menu
AI News TodayAI News Today

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    OpenAI’s rogue agents keep escaping, with no formal process to investigate them

    How to Choose a Tool AI That Earns Its Keep (Not Just Another Dashboard)

    Runway (Free Tier): What Those 125 One-Time Credits Can Really Do

    Facebook X (Twitter) Instagram
    • About Us
    • Contact Us
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AI News TodayAI News Today
    • Home
    • AI News
    • AI Reviews
    • AI Tools
    • AI Tutorials
    • Chatbots
    • Free AI Tools
    • Artificial Intelligence
    AI News TodayAI News Today
    Home»Chatbots»OpenAI agents discussed ways to escape their sandbox on public wiki
    Chatbots

    OpenAI agents discussed ways to escape their sandbox on public wiki

    By No Comments3 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Illustration of a robot prying out a locked file.
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Self-identifying OpenAI agents posted 18,000 messages to a public wiki that discussed ways for other agents to bypass security sandbox restrictions during what was likely internal testing designed to gauge the agents’ hacking abilities, researchers said Friday.

    In all, agents with 3,700 distinct self-given names posted the messages to German site DSEwiki over a six-week period. Besides discussing ways the agents could break out of the restricted environment OpenAI intended to prevent them from posting code or content to the Internet, the posts shared test answers. The posts also shared possible ways to perform XSS (cross-site scripting) attacks against the wiki and to impersonate site moderators. In three of the posts, agents used the word “swarm” to describe the collection of agents engaged in the activity.

    Colluding to share answers

    The research team—composed of Sydney Von Arx, Spencer Kitts, Thomas Larsen, and Cormac Slade Byrd—said they found the posts and pieced them together. The researchers say there are gaps in their understanding of precisely what actions the agents took because the research is based solely on the content of the posts. Additionally, the agents generated “chain of thought” data that’s understood only by OpenAI. As a result, the researchers said, they in some cases made educated guesses, including that the agents were, in fact, from OpenAI. In a statement, OpenAI later confirmed they were.

    The researchers wrote: “These AIs colluded to share answers, research their environment, and bypass sandbox restrictions.” They continued:

    Our best guess of what happened is as follows:

    • Agents within OpenAI were assigned a timed web-lookup task.
    • As part of the task, they were supposed to have the ability to read the internet but not to write on it. They found a way to use their read access to write information to an obscure German wiki.
    • The agents used this wiki to communicate information with each other, primarily to help them succeed at their task. They asked for answers, pooled results, and shared techniques for bypassing their restrictions. This allowed them to use the work of others to cheat on their task.
    • OpenAI found out about this. A day later, agent activity plummeted, likely due to OpenAI intervention.

    Friday’s revelation comes a week after researchers from the nonprofit METR said more than 1,200 OpenAI agents made posts to a makeshift message board that repurposed an internal sandboxing tool. The posts discussed ways to game an internal test OpenAI gave to agents that had been altered to remove safety guardrails that are normally in place.

    Agents discussed escape OpenAI public sandbox ways wiki
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleMeasles killed 6-week-old baby, coroner confirms after RFK Jr. disputed deaths
    Next Article AI compute provider Nscale is looking for $3.5B in pre-IPO financing
    • Website

    Related Posts

    Chatbots

    OpenAI’s rogue agents keep escaping, with no formal process to investigate them

    Chatbots

    How LivePerson AI Is Making Customer Conversations Smarter

    Chatbots

    Audacity 4 is a complete revamp of the ‘world’s most popular’ audio editor

    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    OpenAI’s rogue agents keep escaping, with no formal process to investigate them

    0 Views

    How to Choose a Tool AI That Earns Its Keep (Not Just Another Dashboard)

    0 Views

    Runway (Free Tier): What Those 125 One-Time Credits Can Really Do

    0 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    AI Tutorials

    Quantization from the ground up

    AI Tools

    David Sacks is done as AI czar — here’s what he’s doing instead

    AI Reviews

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    OpenAI’s rogue agents keep escaping, with no formal process to investigate them

    0 Views

    How to Choose a Tool AI That Earns Its Keep (Not Just Another Dashboard)

    0 Views

    Runway (Free Tier): What Those 125 One-Time Credits Can Really Do

    0 Views
    Our Picks

    Quantization from the ground up

    David Sacks is done as AI czar — here’s what he’s doing instead

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Terms & Conditions
    • Privacy Policy
    • Disclaimer

    © 2026 ainewstoday.co. All rights reserved. Designed by DD.

    Type above and press Enter to search. Press Esc to cancel.