Research Article

Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries

Volume: 12 Number: 1 June 30, 2026
EN

Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries

Abstract

This study aimed to evaluate the information generation performance of four large language models (ChatGPT-o1-pro, ChatGPT-4o, Gemini 2.5 Flash, and Grok-3) using multidimensional criteria, based on their responses to expert-level questions in the field of intraoral bone augmentation. In addition, the study investigated the effect of enabling Deep Search mode on model performance. A total of 20 expert-generated questions, developed with input from domain specialists, were submitted to each model under two different retrieval configurations, including Deep Search. The responses were assessed independently by two periodontology experts in terms of accuracy, content quality, readability, and reading level. The findings indicated that the ChatGPT-o1-pro model with Deep Search achieved the highest and most balanced performance across all evaluation criteria. Similarly, ChatGPT-4o with Deep Search demonstrated strong performance, particularly in accuracy and content quality, although its readability was comparatively lower. In contrast, Gemini 2.5 Flash and Grok-3 showed relatively weaker performance across all dimensions. Overall, the results revealed that activating Deep Search mode provided a statistically significant improvement in performance. These findings underscore the growing potential of large language models in clinically complex and information-intensive domains such as dentistry. The results also suggest that, beyond model selection, the operational mode—particularly the use of advanced retrieval configurations—plays a crucial role in determining the quality of generated information.

Keywords

References

  1. Najeeb M, Islam S. Artificial intelligence in restorative dentistry: current trends and future prospects. BMC Oral Health. 2025; 25:592.
  2. Lee SJ, Poon J, Jindarojanakul A, et al. Artificial intelligence in dentistry: exploring emerging applications and future prospects. J Dent. 2025; 155:105648.
  3. Giannakopoulos K, Kavadella A, Salim AA, et al. Evaluation of the performance of generative AI large language models ChatGPT, Google Bard, and Microsoft Bing Chat in supporting evidence-based dentistry: comparative mixed methods study. J Med Internet Res. 2023;25: e51580.
  4. Dermata A, Arhakis A, Makrygiannakis MA, et al. Evaluating the evidence-based potential of six large language models in paediatric dentistry: a comparative study on generative artificial intelligence. Eur Arch Paediatr Dent. 2025; 26:527–35.
  5. Macrì M, D'Albis V, D'Albis G, et al. The role and applications of artificial intelligence in dental implant planning: a systematic review. Bioengineering. 2024; 11:778.
  6. Mukherjee J, Sharma R, Dutta P, Bhunia B. Artificial intelligence in healthcare: a mastery. Biotechnol Genet Eng Rev. 2024; 40:1659–708.
  7. Aftab W, Apostolou Z, Bouazoune K, Straub T. Optimizing biomedical information retrieval with a keyword frequency-driven prompt enhancement strategy. BMC Bioinformatics. 2024; 25:281.
  8. Maslej MM, Donner K, Thakur A, et al. Deriving insights from open-ended learner feedback: an exploration of natural language processing approaches. J Contin Educ Health Prof. 2025.

Details

Primary Language

English

Subjects

Active Sensing

Journal Section

Research Article

Publication Date

June 30, 2026

Submission Date

January 14, 2026

Acceptance Date

May 8, 2026

Published in Issue

Year 2026 Volume: 12 Number: 1

APA
Bayırlı, A. B., Uytun, M., Erdem, R., & Genç, Y. S. (2026). Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries. Mugla Journal of Science and Technology, 12(1), 85-100. https://doi.org/10.22531/muglajsci.1862743
AMA
1.Bayırlı AB, Uytun M, Erdem R, Genç YS. Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries. Mugla Journal of Science and Technology. 2026;12(1):85-100. doi:10.22531/muglajsci.1862743
Chicago
Bayırlı, Ali Batuhan, Mehmetcan Uytun, Ruşen Erdem, and Yavuz Selim Genç. 2026. “Comparative Evaluation of Chatgpt, Gemini and Grok With and Without Deep Research Mode in Answering Bone Augmentation Queries”. Mugla Journal of Science and Technology 12 (1): 85-100. https://doi.org/10.22531/muglajsci.1862743.
EndNote
Bayırlı AB, Uytun M, Erdem R, Genç YS (June 1, 2026) Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries. Mugla Journal of Science and Technology 12 1 85–100.
IEEE
[1]A. B. Bayırlı, M. Uytun, R. Erdem, and Y. S. Genç, “Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries”, Mugla Journal of Science and Technology, vol. 12, no. 1, pp. 85–100, June 2026, doi: 10.22531/muglajsci.1862743.
ISNAD
Bayırlı, Ali Batuhan - Uytun, Mehmetcan - Erdem, Ruşen - Genç, Yavuz Selim. “Comparative Evaluation of Chatgpt, Gemini and Grok With and Without Deep Research Mode in Answering Bone Augmentation Queries”. Mugla Journal of Science and Technology 12/1 (June 1, 2026): 85-100. https://doi.org/10.22531/muglajsci.1862743.
JAMA
1.Bayırlı AB, Uytun M, Erdem R, Genç YS. Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries. Mugla Journal of Science and Technology. 2026;12:85–100.
MLA
Bayırlı, Ali Batuhan, et al. “Comparative Evaluation of Chatgpt, Gemini and Grok With and Without Deep Research Mode in Answering Bone Augmentation Queries”. Mugla Journal of Science and Technology, vol. 12, no. 1, June 2026, pp. 85-100, doi:10.22531/muglajsci.1862743.
Vancouver
1.Ali Batuhan Bayırlı, Mehmetcan Uytun, Ruşen Erdem, Yavuz Selim Genç. Comparative Evaluation of Chatgpt, Gemini and Grok with and without Deep Research Mode in Answering Bone Augmentation Queries. Mugla Journal of Science and Technology. 2026 Jun. 1;12(1):85-100. doi:10.22531/muglajsci.1862743

8805

Mugla Journal of Science and Technology (MJST) is licensed under the Creative Commons Attribution-Noncommercial-Pseudonymity License 4.0 international license