Can ChatGPT Diagnose Your Condition? Not yet

ChatGPT, a sophisticated chatbot driven by artificial intelligence (AI) technology, has been increasingly used in health care contexts, one of which is assisting patients in self-diagnosing before seeking medical help. Although it seems very useful at first glance, AI may cause more harm than good to the patient if it is not accurate in its diagnosis and recommendations. A research team from Japan and the United States recently found that the precision of ChatGPT’s diagnoses and the degree to which it recommends medical consultation require further development.

In a study published in September, the multi-institutional research team led by Tokyo Medical and Dental University (TMDU) evaluated the accuracy (percentage of correct responses) and precision of ChatGPT’s response to five common orthopedic diseases (including carpal tunnel syndrome, cervical myelopathy, and hip osteoarthritis) because orthopedic complaints are very common in clinical practice and comprise up to 26% of the reasons why patients seek care. Over a 5-day course, each of the study researchers submitted the same questions to ChatGPT. The reproducibility between days and researchers was also calculated, and the strength of the recommendation that the patient seek medical attention was evaluated.

"We found that accuracy and reproducibility of ChatGPT's diagnosis are not consistent over the five conditions. ChatGPT's diagnosis was 100% accurate for carpal tunnel syndrome, but only 4% for cervical myelopathy," says lead author Tomoyuki Kuroiwa. Additionally, reproducibility between days and researchers varied from "poor" to "almost perfect" among the five conditions even though researchers entered the same questions every time.

ChatGPT was also inconsistent in recommending medical consultation. Although almost 80% of ChatGPT's answers recommended medical consultation, only 12.8% included a strong recommendation as set by the study standards. "Without direct language, it is possible that the patient is left confused after self-diagnosis, or worse, experience harm from a misdiagnosis," says Kuroiwa.

This is the first study to evaluate the reproducibility and degree of the medical consultation recommendation of ChatGPT’s ability to self-diagnose. "In its current form, ChatGPT is inconsistent in both accuracy and precision to help patients diagnose their disease," explains senior author Koji Fujita. "Given the risk of error and potential harm from misdiagnosis, it is important for any diagnostic tool to include clear language alerting patients to seek expert medical opinions for confirmation of a disease."

The researchers also note some limitations of the study including the use of questions simulated by the research team and not patient-derived questions; focusing on only five orthopedic diseases; and using only ChatGPT. While it is still too early to use AI intelligence for self-diagnosis, the training of ChatGPT on diseases of interest could change this. Future studies can help shed light on the role of AI as a diagnostic tool.

Kuroiwa T, Sarcon A, Ibara T, Yamada E, Yamamoto A, Tsukamoto K, Fujita K.
The Potential of ChatGPT as a Self-Diagnostic Tool in Common Orthopedic Diseases: Exploratory Study.
J Med Internet Res 2023;25:e47621. doi: 10.2196/47621

Most Popular Now

Philips Foundation 2024 Annual Report: E…

Marking its tenth anniversary, Philips Foundation released its 2024 Annual Report, highlighting a year in which the Philips Foundation helped provide access to quality healthcare for 46.5 million people around...

New AI Transforms Radiology with Speed, …

A first-of-its-kind generative AI system, developed in-house at Northwestern Medicine, is revolutionizing radiology - boosting productivity, identifying life-threatening conditions in milliseconds and offering a breakthrough solution to the global radiologist...

Scientists Argue for More FDA Oversight …

An agile, transparent, and ethics-driven oversight system is needed for the U.S. Food and Drug Administration (FDA) to balance innovation with patient safety when it comes to artificial intelligence-driven medical...

New Research Finds Specific Learning Str…

If data used to train artificial intelligence models for medical applications, such as hospitals across the Greater Toronto Area, differs from the real-world data, it could lead to patient harm...

Giving Doctors an AI-Powered Head Start …

Detection of melanoma and a range of other skin diseases will be faster and more accurate with a new artificial intelligence (AI) powered tool that analyses multiple imaging types simultaneously...

AI Agents for Oncology

Clinical decision-making in oncology is challenging and requires the analysis of various data types - from medical imaging and genetic information to patient records and treatment guidelines. To effectively support...

Patients say "Yes..ish" to the…

As artificial intelligence (AI) continues to be integrated in healthcare, a new multinational study involving Aarhus University sheds light on how dental patients really feel about its growing role in...

Brains vs. Bytes: Study Compares Diagnos…

A University of Maine study compared how well artificial intelligence (AI) models and human clinicians handled complex or sensitive medical cases. The study published in the Journal of Health Organization...

'AI Scientist' Suggests Combin…

An 'AI scientist', working in collaboration with human scientists, has found that combinations of cheap and safe drugs - used to treat conditions such as high cholesterol and alcohol dependence...

Start-ups in the Spotlight at MEDICA 202…

17 - 20 November 2025, Düsseldorf, Germany. MEDICA, the leading international trade fair and platform for healthcare innovations, will once again confirm its position as the world's number one hotspot for...