inklap

ChatGPT-4 Omni’s Accuracy in Multiple-Choice Dentistry Questions: A Multidisciplinary and Bilingual Assessment

Makbule Buse Dündar Sarı, Berkant Sezer · Essentials of Dentistry · 2025

Background: This study evaluated the performance of ChatGPT-4 Omni (ChatGPT-4o) in answering multiple-choice questions from the Dentistry Specialty Examination (DUS), a nationwide exam conducted in Türkiye, assessing knowledge in basic medical and clinical dentistry sciences. Additionally, it examined performance variations based on question language (Turkish vs. English). Methods: The dataset included 1504 unique questions from publicly available DUS exams (2012-2021) categorized into Basic Medical Sciences (n= 514) and Clinical Dentistry Sciences (n= 990). Each question was presented to ChatGPT-4o in both Turkish and English, generating 3008 responses. Accuracy was determined using the official answer key. McNemar’s test compared accuracy between languages, while chi-square and Bonferroni post-hoc tests assessed differences across disciplines. Results: ChatGPT-4o showed significantly higher accuracy for English questions (87.8%) than Turkish questions (84.0%) (P < .001). In Basic Medical Sciences, accuracy was significantly higher for English questions in Anatomy (P = .004) and Physiology (P = .039), while Biochemistry achieved 100% accuracy in both languages. In Clinical De

📖 افتح في inklap 🔗 DOI 📮 اطلب بحثاً