Synapse
⌘+K
Synapse
PulseExploreClubsResearchersJournals
Instagram
HomeClubsExplore
April 30, 2026BMC Oral HealthOpen Access

Comparative analysis of the performance of artificial intelligence language models on Turkish dental specialty examination questions

View Full Paper
Ask AI
Bookmark
Share

Authors

GCGökhan CemNYNeslihan YılmazDYDoğukan Yılmaz

Discussion

Loading...

Member takes

Overview

Analysis compares performance of AI language models on dental exam questions, revealing distinctions in multiple-choice answers.

Key Points

  • To compare the performance of AI language models on Turkish dental specialty examination questions.
  • Three AI models evaluated: ChatGPT-4o, Gemini 2.0 Pro, DeepSeek-R1.
  • 1506 DUS questions analyzed through zero-shot prompting.
  • Statistical analyses included Cochran’s Q test and Bonferroni-corrected McNemar tests.
  • Gemini 2.0 Pro scored significantly higher than ChatGPT-4o and DeepSeek-R1 in total correct answers.
  • No significant difference in basic sciences performance was found.
  • Clinical sciences analysis showed significant differences, favoring Gemini 2.0 Pro.

Cite This Study

Cem et al. (2026) studied this question.

synapsesocial.com/papers/69f2a49d8c0f03fd67763afdhttps://doi.org/10.1186/s12903-026-08490-5
View Full Paper
Ask AI
Bookmark
Share

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Evaluating the Accuracy of Large Language Models in Dentistry: A Multi‐Model Study Using Clinical Questions From Turkey's Dental Specialty Exams2026
  2. 2A comparative analysis of the performance of leading large language models on the endodontics section of the dentistry specialization exam in Türkiye2026
  3. 3Performance of Large Language Models on Official Periodontology Questions: A 13-Year Analysis of the Turkish Dental Specialization Examination2025
  4. 4Benchmarking large language models on Turkish dental specialty examination questions: effects of model, discipline, and question format2026
  5. 5Comparative Evaluation of Four Large Language Models in Turkish Dentistry Specialization Exam2025