Drooid Logo
Back to story perspectives

Full Breakdown

Risks of AI Chatbots in Healthcare: A Study Overview

2/10/2026, 4:39:47 PM

Study Findings on AI Chatbot Reliability

A recent study conducted by researchers at the University of Oxford has raised concerns about the reliability of AI chatbots in providing medical advice. The research involved 1,298 participants who were presented with various health scenarios, such as experiencing a severe headache or a new mother feeling exhausted. Participants were divided into groups, with some using AI chatbots—specifically OpenAI's GPT-4o, Meta's Llama 3, and Command R+—while others relied on traditional internet search engines for guidance. The results indicated that individuals using AI chatbots were only able to correctly identify their health issues approximately one-third of the time and determined the appropriate course of action about 45% of the time, which was no better than the control group.

Communication Challenges

The study highlighted significant communication breakdowns between users and AI chatbots. Participants often struggled to provide complete information, which hindered the chatbots' ability to offer accurate advice. Dr. Rebecca Payne, the lead medical practitioner on the study, emphasized that asking chatbots about symptoms could be "dangerous," as they might provide incorrect diagnoses or fail to recognize when urgent medical help is needed. The findings suggest that while AI can perform well on medical benchmarks, its practical application in real-world scenarios remains problematic.

Expert Perspectives

Dr. Adam Mahdi, a senior author of the study, noted that users frequently did not know how to phrase their questions effectively, leading to varied and sometimes confusing responses from the chatbots. He stated, "This is exactly when things would fall apart." Additionally, Dr. Amber W. Childs from Yale School of Medicine pointed out that biases inherent in medical practices could be perpetuated by AI systems trained on existing data.

Dr. Jonathan Chen, an internist at Stanford, expressed concerns about the implications of AI in medicine, stating that these technologies could threaten the identity and purpose of healthcare professionals. He and others in the medical community are grappling with the question of when it might be appropriate to allow AI to take a more prominent role in patient care.

Calls for Improvement and Regulation

In light of the study's findings, experts are advocating for improvements in AI technology, particularly in health-related applications. Dr. Bertalan Meskó, editor of The Medical Futurist, mentioned that recent advancements by AI developers like OpenAI and Anthropic could yield different results in future studies. He stressed the importance of establishing clear national regulations and medical guidelines to ensure the safe use of AI in healthcare.

Conclusion

The Oxford study underscores the potential risks associated with using AI chatbots for medical advice. As reliance on these technologies grows, it is crucial for users to remain cautious and seek information from trusted medical sources, such as the UK's National Health Service. The ongoing development of AI in healthcare necessitates a careful balance between innovation and patient safety.

Verbatim Quotes

  • “Patients need to be aware that asking a large language model about their symptoms can be dangerous, giving wrong diagnoses and failing to recognise when urgent help is needed,” — Dr. Rebecca Payne, Lead Medical Practitioner
  • “They threaten your identity and your purpose.” — Dr. Jonathan Chen, Internist at Stanford