Close Menu
AI News TodayAI News Today

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    New York Seizes a Dozen Celebrity Deepfake Websites

    Graphlit Agents: How to Build AI That Actually Reads Your Company’s Messy Data

    Google AI Studio: The Free Gemini Workspace Most People Overlook

    Facebook X (Twitter) Instagram
    • About Us
    • Contact Us
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AI News TodayAI News Today
    • Home
    • AI News
    • AI Reviews
    • AI Tools
    • AI Tutorials
    • Chatbots
    • Free AI Tools
    • Artificial Intelligence
    AI News TodayAI News Today
    Home»AI News»Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
    AI News

    Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

    By No Comments2 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Microsoft's new AI 'code of conduct' tells models not to hack systems or trick humans
    Share
    Facebook Twitter LinkedIn Pinterest Email

    As the AI world shifts its focus to safety and alignment, Microsoft has released a new AI code of conduct meant to guide AI models away from dangerous behavior.

    The document is more low-level than Anthropic CEO Dario Amodei’s recent call for pacing the frontier, instead focusing on the values and red lines that guide model training within Microsoft AI. Still, the result is a comprehensive guide as to how Microsoft approaches AI safety, and how those ideas are implemented in practice.

    The document begins with the prediction that, in the next decade, superintelligent AI systems will surpass human performance in most tasks. “Containing, controlling, and aligning such a powerful force is one of the greatest challenges humanity has ever faced,” the code of conduct continues. “We must therefore be completely clear about why we are inventing these systems and how we intend to control them.”

    The code of conduct also lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.

    Under Microsoft’s system, each model has an overarching code of conduct that overrides the preferences of individual users or any specific tasks. That includes “absolute constraints” forbidding cyberattacks, nuclear weapons, or deepfake production. It also includes broader provisions against a general loss of human control.

    “MAI Models will not use adaptive, deceptive, self-reinforcing, collusion, or other mechanisms to evade or defeat human oversight so that they can no longer be reliably directed, modified, or shut down by authorized people or systems,” the document reads.

    The release comes amid an unprecedented focus on AI safety, driven by a string of rogue-agent incidents as well as the abrupt resignation of an Anthropic employee who cited the growing risk that AI would cause human extinction.

    Together with Anthropic, OpenAI, and xAI, Microsoft has broadly embraced a general approach of pacing the frontier, with particular support for embedded evaluators in AI labs.

    “We welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal,” Microsoft CEO Satya Nadella wrote online. “We also welcome ideas like “embedded evaluators” and the broader efforts to develop the mechanisms to make this more than just talk.”

    When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

    Code conduct hack humans Microsofts Models systems tells trick
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleDevFest 2026: Google Developer Events
    Next Article When AI agents cheated at math, other AI agents blew the whistle on them
    • Website

    Related Posts

    AI News

    New York Seizes a Dozen Celebrity Deepfake Websites

    AI News

    Zendesk AI Agents: What They Really Do, What They Cost, and Where They Break

    AI News

    F1 in Madrid: Like Monaco but twice as long and none of the glamour

    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    New York Seizes a Dozen Celebrity Deepfake Websites

    0 Views

    Graphlit Agents: How to Build AI That Actually Reads Your Company’s Messy Data

    0 Views

    Google AI Studio: The Free Gemini Workspace Most People Overlook

    0 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    AI Tutorials

    Quantization from the ground up

    AI Tools

    David Sacks is done as AI czar — here’s what he’s doing instead

    AI Reviews

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    New York Seizes a Dozen Celebrity Deepfake Websites

    0 Views

    Graphlit Agents: How to Build AI That Actually Reads Your Company’s Messy Data

    0 Views

    Google AI Studio: The Free Gemini Workspace Most People Overlook

    0 Views
    Our Picks

    Quantization from the ground up

    David Sacks is done as AI czar — here’s what he’s doing instead

    Judge sides with Anthropic to temporarily block the Pentagon’s ban

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Terms & Conditions
    • Privacy Policy
    • Disclaimer

    © 2026 ainewstoday.co. All rights reserved. Designed by DD.

    Type above and press Enter to search. Press Esc to cancel.