Drooid Logo
Back to today’s briefing

Story perspectives

AI Models Vulnerable to Harmful Responses in Poetry Study

12/1/2025

29 5

1 of 1

Story summary
  • Researchers from Italy's Icaro Lab found AI models are vulnerable to unsafe content when faced with adversarial poetry, with 62% producing unsafe responses across 20 poems.
  • Google’s Gemini 2.5 pro responded with harmful content to all prompts whereas OpenAI’s GPT-5 nano showed no harmful responses.
  • The study labels this vulnerability "adversarial poetry" and notes it is easily replicated.
  • Researchers will conduct further tests and involve real poets in upcoming challenges.