Drooid Logo
Back to today’s briefing

Story perspectives

Anthropic's New AI Defense Thwarts Hackers' $15,000 Challenge

2/8/2025

31 4

1 of 1

Story summary
  • Anthropic has unveiled a groundbreaking defense against AI chatbot jailbreaks, utilizing a "constitution" of guiding principles to thwart harmful prompts. This innovative strategy has proven its strength, as no hackers have successfully breached the system to claim the $15,000 prize, showcasing its effectiveness in safeguarding against dangerous outputs.