İçeriğe geç
akaturk Akademik ölçüm

Makale detayı · 2026

Comparative assessment of ChatGPT and Gemini answers to common chronic obstructive pulmonary disease questions: An expert panel evaluation by pulmonologists

Chronic Respiratory Disease

YÖKSİS OpenAlex Açık erişim · gold SJR Q2 JCR Q2 Atıf 0 Yüzdelik 27.8% FWCI 0.0
Yıl
2026
ISSN
1479-9723
Tür
article

Veri kaynağı ayrımı

  • YÖKSİS YÖKSİS makale kaydı
  • OpenAlex OpenAlex zenginleştirmesi (özet, atıf, konular)

Özet

İngilizce (OpenAlex)

Background AI-based chatbots are increasingly used as sources of health information. However, their reliability in delivering accurate and scientifically sound responses to patient questions remains uncertain, especially in chronic diseases such as chronic obstructive pulmonary disease (COPD). This study aims to compare the reliability of ChatGPT-4o and Gemini 2.5 Flash in providing patient-centered medical information on COPD. Methods A total of 34 common public questions about COPD were submitted to ChatGPT-4o and Gemini 2.5 Flash. Responses were evaluated blindly by four pulmonologists across three domains: accuracy, clarity, and scientific adequacy. The mean scores and word counts were analyzed and compared via nonparametric tests. Results Gemini 2.5 Flash outperforms ChatGPT-4o in terms of scientific adequacy (mean score: 4.69 ± 0.31 vs. 4.34 ± 0.45, p <0.001). No significant difference was found in accuracy or clarity. The Gemini 2.5 Flash also generated significantly longer responses, particularly in the treatment and prognosis domains ( p <0.001). Both models provided generally acceptable answers, but ChatGPT-4o′s responses were shorter and occasionally less complete. Conclusions While both models delivered largely accurate and understandable content, Gemini 2.5 Flash tended to produce more detailed responses and received higher scientific adequacy ratings; however, this difference should be interpreted in light of the substantial imbalance in response length. These tools may support patient education however, the findings reflect a comparison between AI systems only and should be interpreted within this scope.

Konular

  • Artificial Intelligence in Healthcare and Education
  • Digital Mental Health Interventions
  • Health Literacy and Information Accessibility

Birincil konu Artificial Intelligence in Healthcare and Education

Yazarlar

  1. MUTLU ONUR GÜÇSAV
  2. DAMLA SERÇE UNAT
  3. ONUR AKÇAY
  4. ÖMER SELİM UNAT
  5. AYSU AYRANCI İZMİR BAKIRÇAY ÜNİVERSİTESİ
  6. AHMET EMİN ERBAYCU