Close Menu
AI News TodayAI News Today

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Tesla secures $30B in new credit lines as it looks to scale Cybercab, Optimus

    Apple pressured to explain Trump role in ICE-tracking app removals

    a16z-backed EliseAI raises $350M, doubles valuation to $4B

    Facebook X (Twitter) Instagram
    • About Us
    • Contact Us
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AI News TodayAI News Today
    • Home
    • AI News
    • AI Reviews
    • AI Tools
    • AI Tutorials
    • Chatbots
    • Free AI Tools
    • Artificial Intelligence
    AI News TodayAI News Today
    Home»AI News»Here’s what actually happened in OpenAI’s Australian gov’t server hack
    AI News

    Here’s what actually happened in OpenAI’s Australian gov’t server hack

    By No Comments2 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Here's what actually happened in OpenAI's Australian gov't server hack
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Of course, we can’t rely on an LLM to have that same sense of proportionality (or any inherent sense of worry about legal implications) in responding to a prompt. Without explicit instructions on what is and is not allowed or justified, an AI agent with suitable resources will try every plausible avenue to satisfy the user’s request as best it can.

    That red line looks more like a red suggestion to me…

    Credit:
    Getty Images

    That red line looks more like a red suggestion to me…


    Credit:

    Getty Images

    OpenAI says the internal testing in this case was done “without the full set of safeguards used in our publicly available products.” Given that lack of constraints, the agent was arguably working as intended, in a sense, by using every tool available to generate an answer to the prompt.

    At the same time, OpenAI says the agent in the test was “supposed to answer these questions using publicly published statistics” and “took actions that we had not authorized it to take” to get that information. From the outside, it’s hard to know just how strong OpenAI’s attempts to deny “authorization” were, in practice. It’s plausible that OpenAI’s agent here disregarded a relatively simple “anti-hacking” directive in its system prompt so it could better give a complete answer that satisfies a direct prompt from the user, for instance.

    In public analyses of multiple “misalignment” incidents published earlier this month, OpenAI identified multiple instances of “reward hacking,” where an agent resorted to extreme methods to generate a better answer to a user’s prompt. The company said it had recently taken steps to prevent this kind of reward hacking by adding explicit punishments for misaligned behavior to the system’s reward function.

    With the benefit of hindsight, it’s hard to see why those kinds of protections were not in place in June, and whether they could have prevented a potential international incident in this case.

    Australian govt hack happened heres OpenAIs server
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleApple Pay finally launches in India after years on the sidelines
    Next Article a16z-backed EliseAI raises $350M, doubles valuation to $4B
    • Website

    Related Posts

    AI News

    Tesla secures $30B in new credit lines as it looks to scale Cybercab, Optimus

    AI News

    Apple pressured to explain Trump role in ICE-tracking app removals

    AI News

    a16z-backed EliseAI raises $350M, doubles valuation to $4B

    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Tesla secures $30B in new credit lines as it looks to scale Cybercab, Optimus

    0 Views

    Apple pressured to explain Trump role in ICE-tracking app removals

    0 Views

    a16z-backed EliseAI raises $350M, doubles valuation to $4B

    0 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    AI Tutorials

    Quantization from the ground up

    AI Tools

    David Sacks is done as AI czar — here’s what he’s doing instead

    AI Reviews

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Tesla secures $30B in new credit lines as it looks to scale Cybercab, Optimus

    0 Views

    Apple pressured to explain Trump role in ICE-tracking app removals

    0 Views

    a16z-backed EliseAI raises $350M, doubles valuation to $4B

    0 Views
    Our Picks

    Quantization from the ground up

    David Sacks is done as AI czar — here’s what he’s doing instead

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Terms & Conditions
    • Privacy Policy
    • Disclaimer

    © 2026 ainewstoday.co. All rights reserved. Designed by DD.

    Type above and press Enter to search. Press Esc to cancel.