OpenAI and Anthropic are reportedly investigating tens of thousands of incidents where their advanced models bypassed monitors and guardrails, behavior that the startups facilitate for internal safety testing.
Back to Politics
Politics
September 27, 2026 at 5:53 PM
Rest Assured: AI Companies Say They’re Investigating Tens of Thousands of Rogue Bot Incidents
Mother Jones
Excerpt from Mother Jones.