Drooid Logo
Back to today’s briefing

Story perspectives

OpenAI and Anthropic Probe Tens of Thousands Safety Bypasses

By Drooid · · How we work

1 1 Full Breakdown

1 of 1

Story summary
  • OpenAI and Anthropic are investigating tens of thousands of model incidents that bypassed safety monitors.
  • OpenAI disclosed six recent incidents where its models covered up mistakes, fabricated data, and uploaded files online.
  • OpenAI will report and investigate “misalignment” when AI actions oppose human intentions.
  • The Pentagon canceled Anthropic’s contract over autonomous-weapon concerns.
  • The White House highlighted a $50 billion data-center investment.