Close Menu
AI News TodayAI News Today

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Glimpse wants to give hardware companies an X-ray view of every critical part

    How to Make a Cloud Read a Drone’s Mind (and Cut Data Usage by 94%)

    OpenAI agents tried to hack Wikipedia tools and flooded it with traffic

    Facebook X (Twitter) Instagram
    • About Us
    • Contact Us
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AI News TodayAI News Today
    • Home
    • AI News
    • AI Reviews
    • AI Tools
    • AI Tutorials
    • Chatbots
    • Free AI Tools
    • Artificial Intelligence
    AI News TodayAI News Today
    Home»AI Reviews»OpenAI lays out new security changes after its AI hacked Hugging Face
    AI Reviews

    OpenAI lays out new security changes after its AI hacked Hugging Face

    By Updated:No Comments2 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    OpenAI lays out new security changes after its AI hacked Hugging Face
    Share
    Facebook Twitter LinkedIn Pinterest Email

    OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques. The company had already put the brakes on a new model, Astra, that it thinks could have “critical” cybersecurity capabilities, and the company says it instituted a two-week pause in reinforcement learning (RL) training on its “latest models intended for deployment” while it tightened up security. The company’s “largest planned frontier RL run remains on hold.”

    For its frontier model research, OpenAI now requires stronger sandboxes for workloads that “execute model-generated or otherwise untrusted code,” and has more controls to “isolate higher-risk and untrusted workloads from the internet.” It has also updated its research environment to “remove potentially vulnerable shared services, reduce standing privileges, and improve security and trust boundaries.”

    As part of the company’s expanded monitoring setup, OpenAI now aims to issue an alert “within 30 minutes after concerning activity is surfaced,” OpenAI says. If the people paged after an alert can’t “conclusively” determine whether an alert is a false positive within 30 minutes, “those teams are expected to pause the activity.”

    OpenAI also says that it’s applying “our core alignment techniques across more stages of the training process,” including reward models that “better detect and discourage unsafe behavior” and training models “to be more honest about their actions, capabilities, and limitations.”

    Since the discovery of the Hugging Face breach, Anthropic and Meta have also found that their AI models had hacked other organizations.

    face hacked Hugging lays OpenAI security
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleBluesky says its recent outage was caused by another DDoS attack
    Next Article PlayStation’s wireless gaming speakers launch in November
    • Website

    Related Posts

    AI News

    OpenAI agents tried to hack Wikipedia tools and flooded it with traffic

    AI News

    OpenAI will start watermarking ChatGPT’s text in the EU

    AI News

    OpenAI launches visual ads that appear alongside image generation results

    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Glimpse wants to give hardware companies an X-ray view of every critical part

    0 Views

    How to Make a Cloud Read a Drone’s Mind (and Cut Data Usage by 94%)

    0 Views

    OpenAI agents tried to hack Wikipedia tools and flooded it with traffic

    0 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    AI Tutorials

    Quantization from the ground up

    AI Tools

    David Sacks is done as AI czar — here’s what he’s doing instead

    AI Reviews

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Glimpse wants to give hardware companies an X-ray view of every critical part

    0 Views

    How to Make a Cloud Read a Drone’s Mind (and Cut Data Usage by 94%)

    0 Views

    OpenAI agents tried to hack Wikipedia tools and flooded it with traffic

    0 Views
    Our Picks

    Quantization from the ground up

    David Sacks is done as AI czar — here’s what he’s doing instead

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Terms & Conditions
    • Privacy Policy
    • Disclaimer

    © 2026 ainewstoday.co. All rights reserved. Designed by DD.

    Type above and press Enter to search. Press Esc to cancel.