Research Article

Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions

Number: Advanced Online Publication Early Pub Date: July 31, 2026

Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions

Abstract

Objective: Large language models are increasingly used to obtain medical information; however, their accuracy in asthma-related board-review questions, particularly those requiring guideline-based clinical reasoning, remains uncertain. This study aimed to compare the performance of different ChatGPT model configurations in asthma-related multiple-choice questions.

Methods: Forty asthma-related questions obtained from an educational board-review resource were submitted separately to seven ChatGPT model configurations in May 2026: GPT-5.2 Instant, GPT-5.2 Standard Thinking, GPT-5.2 Extended Thinking, OpenAI o3, GPT-5.5 Instant, GPT-5.5 Standard Thinking, and GPT-5.5 Extended Thinking. Each question was entered individually in a new temporary chat session. Responses were compared with reference answers reviewed by adult and pediatric allergist-immunologists according to current guideline recommendations and relevant literature. Accuracy was compared using Cochran’s Q test, McNemar’s test, and the Wilcoxon signed-rank test.

Results: Model accuracy ranged from 85.0% to 92.5%. GPT-5.5 Standard Thinking achieved the highest accuracy, with 37 correct responses out of 40, whereas GPT-5.2 Standard Thinking had the lowest accuracy, with 34 correct responses. No statistically significant difference was observed among the seven configurations (p = 0.638). Family-level accuracy was numerically higher for GPT-5.5 than for GPT-5.2, but this difference was not statistically significant (90.8% vs. 86.7%; p = 0.276). Pairwise analyses also showed no statistically significant differences between model comparisons.

Conclusion: ChatGPT models showed generally high accuracy in asthma-related questions, although errors occurred in selected guideline- and context-dependent scenarios. ChatGPT may serve as a supportive tool for medical education and clinical information seeking, but its outputs require expert verification.

Keywords

Ethical Statement

As this study was based exclusively on the evaluation of publicly available educational questions and ChatGPT-generated responses, and did not involve human participants, real patient data, patient records, biological samples, or personally identifiable information, ethics committee approval was not required.

References

  1. Global Initiative for Asthma. Global strategy for asthma management and prevention: 2026 update [Internet]. Fontana (WI): Global Initiative for Asthma; 2026. Available from: https://ginasthma.org/2026-gina-strategy-report.
  2. Guo D, Xiao N, Chen L, Zhang S. Artificial intelligence in allergic diseases prediction and management. Curr Treat Options Allergy. 2026;13:7. doi:10.1007/s40521-026-00413-7.
  3. Khan M, Banerjee S, Muskawad S, Maity R, Chowdhury SR, Ejaz R, et al. The impact of artificial intelligence on allergy diagnosis and treatment. Curr Allergy Asthma Rep. 2024;24:361-72. doi:10.1007/s11882-024-01152-y.
  4. Cilluffo G, Bajowala S, Elliott T, Gupta P, Volertas S, Ferrante G. The future of allergy management: how artificial intelligence is changing the game. J Allergy Clin Immunol Pract. 2025;13(11):2877-89. doi:10.1016/j.jaip.2025.08.033.
  5. Seurig S, Traidl S, Mathes S. The role of artificial intelligence in modern allergology: a review of applications in diagnosis, prediction, and management. JEADV Clin Pract. 2026:e70297. doi:10.1002/jvc2.70297.
  6. Goktas P, Karakaya G, Kalyoncu AF, Damadoglu E. Artificial intelligence chatbots in allergy and immunology practice: where have we been and where are we going? J Allergy Clin Immunol Pract. 2023;11(9):2697-700. doi:10.1016/j.jaip.2023.05.042.
  7. van Breugel M, Greenhawt M, Eguiluz-Gracia I, Torres Jaén MJ, Anagnostou A, Koppelman GH. Artificial intelligence in allergy and immunology: recent developments, implementation challenges, and the road toward clinical impact. J Allergy Clin Immunol. 2026;157(2):295-306. doi:10.1016/j.jaci.2025.08.022.
  8. Rao AS, Esmail KP, Lee RS, Jiang S, Arraiza Carlo B, Gill J, et al. Large language model performance and clinical reasoning tasks. JAMA Netw Open. 2026;9(4):e264003. doi:10.1001/jamanetworkopen.2026.4003.

Details

Primary Language

English

Subjects

Clinical Sciences (Other)

Journal Section

Research Article

Early Pub Date

July 31, 2026

Publication Date

-

Submission Date

June 2, 2026

Acceptance Date

July 4, 2026

Published in Issue

Year 2026 Number: Advanced Online Publication

APA
Dindar Çelik, F., & Çelik, E. (2026). Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions. Cerasus Journal of Medicine, Advanced Online Publication. https://izlik.org/JA82YZ52RU
AMA
1.Dindar Çelik F, Çelik E. Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions. Cerasus J Med. 2026;(Advanced Online Publication). https://izlik.org/JA82YZ52RU
Chicago
Dindar Çelik, Fatma, and Enes Çelik. 2026. “Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions”. Cerasus Journal of Medicine, no. Advanced Online Publication. https://izlik.org/JA82YZ52RU.
EndNote
Dindar Çelik F, Çelik E (July 1, 2026) Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions. Cerasus Journal of Medicine Advanced Online Publication
IEEE
[1]F. Dindar Çelik and E. Çelik, “Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions”, Cerasus J Med, no. Advanced Online Publication, July 2026, [Online]. Available: https://izlik.org/JA82YZ52RU
ISNAD
Dindar Çelik, Fatma - Çelik, Enes. “Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions”. Cerasus Journal of Medicine. Advanced Online Publication (July 1, 2026). https://izlik.org/JA82YZ52RU.
JAMA
1.Dindar Çelik F, Çelik E. Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions. Cerasus J Med. 2026. Available at https://izlik.org/JA82YZ52RU.
MLA
Dindar Çelik, Fatma, and Enes Çelik. “Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions”. Cerasus Journal of Medicine, no. Advanced Online Publication, July 2026, https://izlik.org/JA82YZ52RU.
Vancouver
1.Fatma Dindar Çelik, Enes Çelik. Comparative Accuracy of Different ChatGPT Model Configurations in Asthma-Related Questions. Cerasus J Med [Internet]. 2026 Jul. 1;(Advanced Online Publication). Available from: https://izlik.org/JA82YZ52RU

INDEXING & ABSTRACTING & ARCHIVING



  Cerasus Journal of Medicine, hosted by Dergipark, is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.