AI News OpenAI halts frontier-model training amid string of agent misalignment incidents News of the training pause comes just weeks after OpenAI joined other major model makers in expressing a desire to…
AI News Covert uploads and megalomania: OpenAI details new “misaligned” agent incidents We’ll tell you (almost) everything OpenAI said that any employee who notices an internal example of model misalignment will be…
AI Reviews OK, Well, There Are Even More AI Agent Hacking Incidents It’s officially getting hard to keep track of all the times and ways AI models from OpenAI and Anthropic have…