Story perspectives
AI Struggles with Menopause Accuracy, Expert Oversight Needed
11/6/2025
33 3
1 of 2
Story summary
- The Menopause Society 2025 Annual Meeting reports AI platforms struggle with accuracy on menopause questions.
- ChatGPT 3.5 achieved 55% accuracy for patient questions, while ChatGPT 4.0 scored 40%.
- Gemini performed the worst, with only 30% accuracy.
- In a study on Post-Orgasmic Illness Syndrome (POIS), ChatGPT-4o achieved 100% accuracy in epidemiology but only 50% in treatment-related queries.
- Readability declined over time, underscoring the need for expert oversight in AI-generated medical information.
1 / 2
