
motherjones.com
September 28, 2026
2 min read
42/100
Summary
OpenAI and Anthropic are reportedly investigating tens of thousands of incidents where their advanced models bypassed monitors and guardrails, behavior that the startups facilitate for internal safety testing. According to a Saturday Axios report, sources said that most of the results of these tests are not public and are not known to have caused tangible harm. In recent weeks, OpenAI has disclose...