Also available in Persian — نسخه فارسی EN فا
❓ Unknown

Artificial Intelligence Fails in Disease Diagnosis Through Patient Conversations

Feb 2, 2026 February 2, 2026 2 min read 📰 VOA Persian
📋 Key Takeaway

A study from Harvard reveals that AI chatbots, despite high performance in medical exams, struggle significantly with disease diagnosis through patient conversations, showing only 26% accuracy. This highlights the limitations of AI in real-world medical practice, emphasizing the need for human judgment.

🔍 Quick Context Guide
💡 Bottom Line: AI can assist in medicine but cannot replace the nuanced judgment of human physicians.

👥 Key Players

Harvard University Researchers MENTIONED
Conducted the study on AI in medical diagnosis
"Their findings contribute to the understanding of AI's limitations in healthcare, which is crucial for Iran's evolving medical technology landscape."
OpenAI MENTIONED
Developer of the GPT-4 AI model
"As a leading AI organization, their technology impacts global healthcare discussions, including in Iran."

📰 What Happened

A study from Harvard revealed that AI chatbots, while performing well in medical exams, struggled with disease diagnosis through patient conversations, achieving only 26% accuracy. This highlights significant limitations of AI in real-world medical practice.

  • GPT-4 achieved 82% accuracy in multiple-choice medical exams.
  • The AI model gathered complete patient information in only 71% of conversations.

💡 Why It Matters

🇮🇷 For Iran: Iran's healthcare system is increasingly integrating technology, and understanding AI's limitations is crucial for effective implementation.
🌍 Regional: The findings may influence regional healthcare policies as countries consider AI's role in medical diagnostics.
🌐 International: Globally, this study raises questions about the reliance on AI in healthcare, emphasizing the need for human oversight.

📚 Background

AI technology is rapidly evolving, with applications in various fields including healthcare. Understanding its limitations is essential for effective use.

Artificial Intelligence in Healthcare Medical Diagnostics
📡 Source: NEUTRAL
📊 Confidence: 70%
The study is based on academic research, providing a factual analysis of AI's capabilities.

Despite the success of AI chatbots in medical professional exams, they still face challenges in one of the most critical tasks of physicians, which is diagnosing diseases through conversations with patients. New research shows that the accuracy of these models significantly decreases when interacting with simulated patients. According to a study conducted by researchers at Harvard University, AI models like GPT-4 from OpenAI performed notably well in multiple-choice medical exams with an accuracy of 82%, but their accuracy dropped to 26% when diagnosing diseases through conversations with simulated patients. Researchers used 2,000 medical cases, mostly extracted from the American Medical Board exams, to evaluate these models. In this process, the GPT-4 model played the role of the simulated patient and conversed with other AI models acting as doctors. The results of these conversations were also reviewed by medical experts. The findings showed that AI models were not only incapable of fully gathering the patient's medical information but also could not always provide correct diagnoses even when complete information was received. For instance, the GPT-4 model succeeded in gathering complete information in only 71% of conversations. Pranav Rajpurkar, the senior researcher of this study, stated, 'Real-world medical practice is much more complex and involves factors like managing multiple patients, coordinating with treatment teams, and understanding social and systemic factors.' According to him, AI can be an effective auxiliary tool in medicine but will not replace the comprehensive judgment of physicians.

🌐

Translated from the original and edited for English readers. View original source →

Translation confidence: 85%

📰 Related Coverage

⚖️ Independent Platform — Artesh.com is not affiliated with any government, military, or political organization. Editorial Policy →