Authors
Shuma Hamaguchi, Masakazu Hamada, Shunya Ikeda, Satoru Kusaka, Tatsuya Akitomo, Ryota Nomura
Published in
Journal of dental sciences. Volume 21. Issue 2. Pages 833-837. Epub Apr 01, 2026.
Abstract
Artificial intelligence (AI) has become widely used and applied in various fields. Although several studies have been conducted using generative AI for various qualification exams, to the best of our knowledge, none has focused on performance changes over time.
In August 2025, ChatGPT 5, Gemini 2.5, Microsoft Copilot, and Medi-Search were asked to answer compulsory questions from five years of the Japanese National Dental Examination. In 2024, we also conducted similar tests on other ChatGPT series and Gemini, and the scores were compared.
In 2025, Copilot, Gemini, and MediSearch scored 80 % or higher, which was the passing standard, for all five years. Although ChatGPT 3.5 did not meet the passing standard for any of the five years, ChatGPT 4o mini and ChatGPT 5 exceeded it for two and three years, respectively. In addition, both Chat GPT's and Gemini's scores substantially improved over time and with each update.
This report suggests that generative AI is improving annually and adapting to the National Dental Examination. Although each AI model is suited to different fields, the trends may change over time. It is necessary to continue comparing and analyzing AI models and provide users with the latest information.
PMID:
42689089
Bibliographic data and abstract were imported from PubMed on 03 Sep 2026.
Read full publication at:
Please sign in
to see all details.
Advertisement
Stats
- Recommendations n/a n/a positive of 0 vote(s)
- Views 6
- Comments 0