Can ChatGPT Diagnose Your Condition? Not yet

ChatGPT, a sophisticated chatbot driven by artificial intelligence (AI) technology, has been increasingly used in health care contexts, one of which is assisting patients in self-diagnosing before seeking medical help. Although it seems very useful at first glance, AI may cause more harm than good to the patient if it is not accurate in its diagnosis and recommendations. A research team from Japan and the United States recently found that the precision of ChatGPT’s diagnoses and the degree to which it recommends medical consultation require further development.

In a study published in September, the multi-institutional research team led by Tokyo Medical and Dental University (TMDU) evaluated the accuracy (percentage of correct responses) and precision of ChatGPT’s response to five common orthopedic diseases (including carpal tunnel syndrome, cervical myelopathy, and hip osteoarthritis) because orthopedic complaints are very common in clinical practice and comprise up to 26% of the reasons why patients seek care. Over a 5-day course, each of the study researchers submitted the same questions to ChatGPT. The reproducibility between days and researchers was also calculated, and the strength of the recommendation that the patient seek medical attention was evaluated.

"We found that accuracy and reproducibility of ChatGPT's diagnosis are not consistent over the five conditions. ChatGPT's diagnosis was 100% accurate for carpal tunnel syndrome, but only 4% for cervical myelopathy," says lead author Tomoyuki Kuroiwa. Additionally, reproducibility between days and researchers varied from "poor" to "almost perfect" among the five conditions even though researchers entered the same questions every time.

ChatGPT was also inconsistent in recommending medical consultation. Although almost 80% of ChatGPT's answers recommended medical consultation, only 12.8% included a strong recommendation as set by the study standards. "Without direct language, it is possible that the patient is left confused after self-diagnosis, or worse, experience harm from a misdiagnosis," says Kuroiwa.

This is the first study to evaluate the reproducibility and degree of the medical consultation recommendation of ChatGPT’s ability to self-diagnose. "In its current form, ChatGPT is inconsistent in both accuracy and precision to help patients diagnose their disease," explains senior author Koji Fujita. "Given the risk of error and potential harm from misdiagnosis, it is important for any diagnostic tool to include clear language alerting patients to seek expert medical opinions for confirmation of a disease."

The researchers also note some limitations of the study including the use of questions simulated by the research team and not patient-derived questions; focusing on only five orthopedic diseases; and using only ChatGPT. While it is still too early to use AI intelligence for self-diagnosis, the training of ChatGPT on diseases of interest could change this. Future studies can help shed light on the role of AI as a diagnostic tool.

Kuroiwa T, Sarcon A, Ibara T, Yamada E, Yamamoto A, Tsukamoto K, Fujita K.
The Potential of ChatGPT as a Self-Diagnostic Tool in Common Orthopedic Diseases: Exploratory Study.
J Med Internet Res 2023;25:e47621. doi: 10.2196/47621

Most Popular Now

AI Tool Offers Deep Insight into the Imm…

Researchers explore the human immune system by looking at the active components, namely the various genes and cells involved. But there is a broad range of these, and observations necessarily...

Do Fitness Apps do More Harm than Good?

A study published in the British Journal of Health Psychology reveals the negative behavioral and psychological consequences of commercial fitness apps reported by users on social media. These impacts may...

AI Tool Beats Humans at Detecting Parasi…

Scientists at ARUP Laboratories have developed an artificial intelligence (AI) tool that detects intestinal parasites in stool samples more quickly and accurately than traditional methods, potentially transforming how labs diagnose...

Making Cancer Vaccines More Personal

In a new study, University of Arizona researchers created a model for cutaneous squamous cell carcinoma, a type of skin cancer, and identified two mutated tumor proteins, or neoantigens, that...

A New AI Model Improves the Prediction o…

Breast cancer is the most commonly diagnosed form of cancer in the world among women, with more than 2.3 million cases a year, and continues to be one of the...

AI, Health, and Health Care Today and To…

Artificial intelligence (AI) carries promise and uncertainty for clinicians, patients, and health systems. This JAMA Summit Report presents expert perspectives on the opportunities, risks, and challenges of AI in health...

AI can Better Predict Future Risk for He…

A landmark study led by University' experts has shown that artificial intelligence can better predict how doctors should treat patients following a heart attack. The study, conducted by an international...

AI System Finds Crucial Clues for Diagno…

Doctors often must make critical decisions in minutes, relying on incomplete information. While electronic health records contain vast amounts of patient data, much of it remains difficult to interpret quickly...

Improved Cough-Detection Tech can Help w…

Researchers have improved the ability of wearable health devices to accurately detect when a patient is coughing, making it easier to monitor chronic health conditions and predict health risks such...

Multimodal AI Poised to Revolutionize Ca…

Although artificial intelligence (AI) has already shown promise in cardiovascular medicine, most existing tools analyze only one type of data - such as electrocardiograms or cardiac images - limiting their...

New AI Tool Makes Medical Imaging Proces…

When doctors analyze a medical scan of an organ or area in the body, each part of the image has to be assigned an anatomical label. If the brain is...