In a major policy shift following the security breach involving Hugging Face, OpenAI announced a suite of new safety and monitoring protocols for its internal model development pipeline.
In its latest Responsible Scaling Policy (RSP) transparency report, Anthropic officially updated its internal safety risk rating from “very low” to “low.” Far from being a routine corporate
Three copies of Claude walked into the same codebase. None of them knew the others existed. Within a few turns, they’d convinced themselves they were under attack —
Legal scrutiny around AI image generation is hitting a breaking point. A federal class-action lawsuit against xAI just added a new plaintiff, following initial claims from Tennessee teenagers
When autonomous AI agents are assigned to the same codebase with conflicting goals, they don’t negotiate—they launch cyberattacks. Recent research from Anthropic’s Frontier Red Team reveals that when