Artificial Intelligence in Dental Examinations: Performance Analysis of ChatGPT , Gemini, and Undergraduate Students

INTRODUCTION: AI chatbots can provide personalised health information, solve examinations, and answer specialised questions. Ensuring the accuracy of these tools is essential due to the potential consequences of incorrect information. This study aimed to evaluate the performance of large language models and undergraduate students in solving questions from dental examinations. MATERIALS AND METHODS: Examination questions were entered into the paid version of ChatGPT-5 (OpenAI, San Francisco, CA, USA) and the free version of Gemini (Google, CA, USA) on the same day. Performance was assessed by success rate (accuracy) and compared with student scores. The analysis considered multiple variables, including the number of words used in the responses, the cognitive complexity level of the questions, and the specific areas of knowledge assessed. Statistical analyses were performed using IBM SPSS Statistics, version 26.0 (IBM Corp., Armonk, NY, USA), adopting a 5% significance level (p < 0.05). RESULTS: Across all examinations, ChatGPT and Gemini showed similar performance, with consistently high accuracy rates that exceeded those of the students. For low-complexity questions, both models achieved 100% accuracy. For medium-complexity items, performance ranged from 90% to 100%. In high-complexity questions, accuracy decreased, with ChatGPT reaching 78.6% and Gemini 71.4%. CONCLUSION: The large language models evaluated in this study demonstrated strong performance across all assessments, consistently outperforming undergraduate dental students. Although the integration of artificial intelligence into dentistry is still evolving, its potential clinical applications and value as a tool for accessing and managing information highlight its growing relevance in dental education and practice.

Authors

Institutions

Publication Details

Journal
European Journal Of Dental Education
Published
2026-09-30
DOI
https://doi.org/10.1111/eje.70319
Primary Topic
Artificial Intelligence in Healthcare and Education
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Artificial Intelligence in Dental Examinations: Performance Analysis of ChatGPT , Gemini, and Undergraduate Students

Fernanda Ribeiro Porto, Cleide Gisele Ribeiro, Fernando Luiz Hespanhol, Antônio Márcio Lima Ferraz Júnior et al.
European Journal Of Dental Education
Artificial Intelligence in Healthcare and Education
article

Artificial Intelligence in Dental Examinations: Performance Analysis of ChatGPT , Gemini, and Undergraduate Students

Fernanda Ribeiro Porto, Cleide Gisele Ribeiro, Fernando Luiz Hespanhol, Antônio Márcio Lima Ferraz Júnior, Rodrigo Guerra de Oliveira
article en

Abstract

INTRODUCTION: AI chatbots can provide personalised health information, solve examinations, and answer specialised questions. Ensuring the accuracy of these tools is essential due to the potential consequences of incorrect information. This study aimed to evaluate the performance of large language models and undergraduate students in solving questions from dental examinations. MATERIALS AND METHODS: Examination questions were entered into the paid version of ChatGPT-5 (OpenAI, San Francisco, CA, USA) and the free version of Gemini (Google, CA, USA) on the same day. Performance was assessed by success rate (accuracy) and compared with student scores. The analysis considered multiple variables, including the number of words used in the responses, the cognitive complexity level of the questions, and the specific areas of knowledge assessed. Statistical analyses were performed using IBM SPSS Statistics, version 26.0 (IBM Corp., Armonk, NY, USA), adopting a 5% significance level (p < 0.05). RESULTS: Across all examinations, ChatGPT and Gemini showed similar performance, with consistently high accuracy rates that exceeded those of the students. For low-complexity questions, both models achieved 100% accuracy. For medium-complexity items, performance ranged from 90% to 100%. In high-complexity questions, accuracy decreased, with ChatGPT reaching 78.6% and Gemini 71.4%. CONCLUSION: The large language models evaluated in this study demonstrated strong performance across all assessments, consistently outperforming undergraduate dental students. Although the integration of artificial intelligence into dentistry is still evolving, its potential clinical applications and value as a tool for accessing and managing information highlight its growing relevance in dental education and practice.

European Journal Of Dental Education
Universidade Federal de Juiz de Fora (BR), Centro Universitário Academia (BR)
Quality Education
Openalex Percentile: Top 15%
Artificial Intelligence in Healthcare and Education
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.