Perween N, Khurana PRS, Aggarwal A, Chaudhary A, Kumar KN, Hassan S, Sekhar AR. Evaluation of ChatGPT’s Accuracy, Repeatability, and Reasoning Ability in Prosthodontics Education: A Cross-Sectional Comparative Study with Prosthodontists. J Clin Exp Dent. 2026;18(3):e342-50.

 

doi:10.4317/jced.63583

https://doi.org/10.4317/jced.63583

___

 

References

1. Schwendicke F, Samek W, Krois J. Artificial Intelligence in Dentistry: Chances and Challenges. J Dent Res. 2020;99(7):769-774.
https://doi.org/10.1177/0022034520915714
PMid:32315260 PMCid:PMC7309354

 

2. Eysenbach G. The Role of ChatGPT, Generative Language Models, and Artificial Intelligence in Medical Education: A Conversation With ChatGPT and a Call for Papers. JMIR Med Educ. 2023;9:e46885.
https://doi.org/10.2196/46885
PMid:36863937 PMCid:PMC10028514

 

3. Shan T, Tay FR, Gu L. Application of Artificial Intelligence in Dentistry. J Dent Res. 2021;100(3):232-244.
https://doi.org/10.1177/0022034520969115
PMid:33118431

 

4. Krusche M, Callhoff J, Knitza J, Ruffer N. Diagnostic accuracy of a large language model in rheumatology: comparison of physician and ChatGPT-4. Rheumatol Int. 2024;44(2):303-306.
https://doi.org/10.1007/s00296-023-05464-6
PMid:37742280 PMCid:PMC10796566

 

5. Antaki F, Touma S, Milad D, El-Khoury J, Duval R. Evaluating the Performance of ChatGPT in Ophthalmology: An Analysis of Its Successes and Shortcomings. Ophthalmol Sci. 2023;3(4):100324.
https://doi.org/10.1016/j.xops.2023.100324
PMid:37334036 PMCid:PMC10272508

 

6. Yeo YH, Samaan JS, Ng WH, Ting PS, Trivedi H, Vipani A, et al. Assessing the performance of ChatGPT in answering questions regarding cirrhosis and hepatocellular carcinoma. Clin Mol Hepatol. 2023;29(3):721-732.
https://doi.org/10.3350/cmh.2023.0089
PMid:36946005 PMCid:PMC10366809

 

7. Revilla-León M, Gómez-Polo M, Vyas S, Barmak AB, Gallucci GO, Att W, et al. Artificial intelligence models for tooth-supported fixed and removable prosthodontics: A systematic review. J Prosthet Dent. 2023;129(2):276-292.
https://doi.org/10.1016/j.prosdent.2021.06.001
PMid:34281697

 

8. Sallam M, Salim NA, Barakat M, Al-Tammemi AB. ChatGPT applications in medical, dental, pharmacy, and public health education: A descriptive study highlighting the advantages and limitations. Narra J. 2023;3(1):e103.
https://doi.org/10.52225/narra.v3i1.103
PMid:38450035 PMCid:PMC10914078

 

9. Shrivastava S, Gupta A, Nanda P, Kaur R, Gupta M, Sharma A, et al. Performance of ChatGPT in dentistry: a cross-sectional, multi-specialty and multi-centric study. Braz J Oral Sci. 2025;24:e254954.
https://doi.org/10.20396/bjos.v24i00.8674954

 

10. Guides to standards in prosthetic dentistry. A report by the British Society for the Study of Prosthetic Dentistry. Br Dent J. 1981;150(6):167-9.
https://doi.org/10.1038/sj.bdj.4804565
PMid:6938213

 

11. Cicchetti DV. Guidelines, criteria, and rules of thumb for evaluating normed and standardized assessment instruments in psychology. Psychol Assess. 1994;6(4):284-290.
https://doi.org/10.1037/1040-3590.6.4.284
PMCid:PMC5872378

 

12. Shrout PE, Fleiss JL. Intraclass correlations: uses in assessing rater reliability. Psychol Bull. 1979;86(2):420-8.
https://doi.org/10.1037/0033-2909.86.2.420
PMid:18839484

 

13. Stratford PW, Goldsmith CH. Use of the standard error as a reliability index of interest: an applied example using elbow flexor strength data. Phys Ther. 1997;77(7):745-50.
https://doi.org/10.1093/ptj/77.7.745
PMid:9225846

 

14. Freire Y, Santamaría Laorden A, Orejas Pérez J, Gómez Sánchez M, Díaz-Flores García V, Suárez A. ChatGPT performance in prosthodontics: Assessment of accuracy and repeatability in answer generation. J Prosthet Dent. 2024;131(4):659.e1-659.e6.
https://doi.org/10.1016/j.prosdent.2024.01.018
PMid:38310063

 

15. Suárez A, Díaz-Flores García V, Algar J, Gómez Sánchez M, Llorente de Pedro M, Freire Y. Unveiling the ChatGPT phenomenon: Evaluating the consistency and accuracy of endodontic question answers. Int Endod J. 2024;57(1):108-113.
https://doi.org/10.1111/iej.13985
PMid:37814369

 

16. Gilson A, Safranek CW, Huang T, Socrates V, Chi L, Taylor RA, et al. How Does ChatGPT Perform on the United States Medical Licensing Examination (USMLE)? The Implications of Large Language Models for Medical Education and Knowledge Assessment. JMIR Med Educ. 2023;9:e45312.
https://doi.org/10.2196/45312
PMid:36753318 PMCid:PMC9947764

 

17. Kahraman F, Aktas A, Bayrakceken S, Çakar T, Tarcan HS, Bayram B, et al. Physicians' ethical concerns about artificial intelligence in medicine: a qualitative study: "The final decision should rest with a human". Front Public Health. 2024;12:1428396.
https://doi.org/10.3389/fpubh.2024.1428396
PMid:39664534 PMCid:PMC11631923

 

18. Umer F, Habib S. Critical Analysis of Artificial Intelligence in Endodontics: A Scoping Review. J Endod. 2022;48(2):152-160.
https://doi.org/10.1016/j.joen.2021.11.007
PMid:34838523

 

19. Kung TH, Cheatham M, Medenilla A, Sillos C, De Leon L, Elepaño C, et al. Performance of ChatGPT on USMLE: Potential for AI-assisted medical education using large language models. PLOS Digit Health. 2023;2(2):e0000198.
https://doi.org/10.1371/journal.pdig.0000198
PMid:36812645 PMCid:PMC9931230

 

20. Liu M, Okuhara T, Huang W, Ogihara A, Nagao HS, Okada H, et al. Large Language Models in Dental Licensing Examinations: Systematic Review and Meta-Analysis. Int Dent J. 2025;75(1):213-222.
https://doi.org/10.1016/j.identj.2024.10.014
PMid:39532572 PMCid:PMC11806300

 

21. Mago J, Sharma M. The Potential Usefulness of ChatGPT in Oral and Maxillofacial Radiology. Cureus. 2023;15(7):e42133.
https://doi.org/10.7759/cureus.42133

 

22. Ahmed WM, Azhari AA, Fawaz KA, Ahmed HM, Alsadah ZM, Majumdar A, et al. Artificial intelligence in the detection and classification of dental caries. J Prosthet Dent. 2025;133(5):1326-1332.
https://doi.org/10.1016/j.prosdent.2023.07.013
PMid:37640607

 

23. Mohammad-Rahimi H, Motamedian SR, Rohban MH, Krois J, Uribe SE, Mahmoudinia E, et al. Deep learning for caries detection: A systematic review. J Dent. 2022;122:104115.
https://doi.org/10.1016/j.jdent.2022.104115
PMid:35367318

 

24. Suárez A, Freire Y, Suárez M, Díaz-Flores García V, Andreu-Vázquez C, Thuissard Vasallo IJ, et al. Diagnostic Performance of Multimodal Large Language Models in the Analysis of Oral Pathology. Oral Dis. 2025;31:70009.
https://doi.org/10.1111/odi.70009
PMid:40545674

 

25. Kıyak YS. ChatGPT's ability or prompt quality: what determines the success of generating multiple-choice questions. Acad Pathol. 2024;11(2):100119.
https://doi.org/10.1016/j.acpath.2024.100119
PMid:38883972 PMCid:PMC11178968

 

26. Freire Y, Santamaría Laorden A, Orejas Pérez J, Ortiz Collado I, Gómez Sánchez M, Thuissard Vasallo IJ, et al. Evaluating the influence of prompt formulation on the reliability and repeatability of ChatGPT in implant-supported prostheses. PLoS One. 2025;20(5):e0323086.
https://doi.org/10.1371/journal.pone.0323086
PMid:40445924 PMCid:PMC12124515

 

27. Alkaissi H, McFarlane SI. Artificial Hallucinations in ChatGPT: Implications in Scientific Writing. Cureus. 2023;15(2):e35179.
https://doi.org/10.7759/cureus.35179

 

28. Jaleel A, Aziz U, Farid G, Zahid Bashir M, Mirza TR, Khizar Abbas SM, et al. Evaluating the Potential and Accuracy of ChatGPT-3.5 and 4.0 in Medical Licensing and In-Training Examinations: Systematic Review and Meta-Analysis. JMIR Med Educ. 2025;11:e68070.
https://doi.org/10.2196/68070
PMid:40973108 PMCid:PMC12495368

 

29. Kochanek K, Skarzynski H, Jedrzejczak WW. Accuracy and Repeatability of ChatGPT Based on a Set of Multiple-Choice Questions on Objective Tests of Hearing. Cureus. 2024;16(5):e59857.
https://doi.org/10.7759/cureus.59857
PMid:38854312 PMCid:PMC11157293