Skip to content
Open access

Comparative Evaluation of Large Language Models as Virtual Orthodontic Patient Information Assistants

Aug 2026 · Journal of International Dental Sciences · 0 citations · 22 references

TL;DR

It is demonstrated that LLMs can serve as auxiliary tools in providing orthodontic patient information but responses should be checked by an expert before being presented to patients and should be adapted into simpler and more understandable language.

Abstract

Objective: The aim of this study was to evaluate the potential use of widely used large language models (LLMs) in patient education by comparing their responses to orthodontic patient questions in terms of accuracy, comprehensiveness, and readability.Methods: This study evaluated the performance of Claude Sonnet 4.6, ChatGPT 5.5 and Gemini 3.5 Pro in providing orthodontic patient information.Ten frequently asked questions by orthodontic patients were posed to each model in independent sessions using the same wording. The responses obtained were scored by two orthodontists for accuracy and comprehensiveness using a modified Likert scale. The readability of the responses was calculated using the FRE and FKGL indices. Differences between the models were evaluated with statistical analyses appropriate for the data structure (p < 0.05).Results: No significant difference was found between the models regarding readability. However, significant differences were found in accuracy and comprehensiveness scores. Gemini 3.5 Pro achieved the highest scores in terms of comprehensiveness and accuracy showed significant superiority over ChatGPT 5.5 (p = 0.007).Conclusion: This study demonstrates that LLMs can serve as auxiliary tools in providing orthodontic patient information.However, responses should be checked by an expert before being presented to patients and should be adapted into simpler and more understandable language.

Read PDF

Similar papers

Open access Sep 2026

Comparative evaluation of large language models for patient-facing information in clear aligner therapy

Large language models can provide appropriate patient-facing information about clear aligner therapy, though meaningful differences in communication quality and empathy exist, though meaningful differences in communication quality and empathy exist.

Atanu Mukhopadhyay, R. Biswas, Sayani Adhikari et al. · 0 citations
Open access Aug 2026

Comparative performance and temporal variability of large language models on orthodontic questions from a national dental specialty examination

This study compared the performance of three artificial intelligence–based chatbots (ChatGPT-4o, ChatGPT-4.5, and Gemini 2.5 Pro) on orthodontic questions from the Turkish Dental Specialty Examination at two testing time points. A total of 179 orthodontic multiple-choice questions from 18 examinations conducted between...

Kübra Arslan Çarpar · 0 citations
#small language model Review Open access Sep 2026

Comparative Evaluation of Large Language Models in Responding to Common Misconceptions Regarding Periodontal Disease and Oral Hygiene

Background/Objective: As large language models (LLMs) are increasingly used to obtain oral health information, their ability to provide accurate and comprehensive responses to patient-focused periodontal questions derived from commonly reported misconceptions is increasingly relevant to patient education. This study ai...

İsmail Gül, Resül Çolak, Merve Küçükoğlu Çolak et al. · 0 citations
Open access Sep 2026

Comparative Evaluation of Large Language Models’ Accuracy in Answering Multiple-Choice Restorative Dentistry Questions From a National Specialty Examination

Objective: Although the integration of large language models (LLMs) into dental education is rapidly increasing, their actual performance in domain-specific assessments remains unclear. This study aimed to evaluate and compare the accuracy of four LLMs (ChatGPT-4.0, Gemini Advanced 1.5 Pro, DeepSeek-V3, and Perplexity)...

Çilem Bulut, Gulben Colak, Gürkan Çolak · 0 citations
Open access Aug 2026

Accuracy and Completeness of Contemporary Large Language Models in Prosthodontics: An Expert-Based Comparative Study

Comparing the scientific accuracy and response completeness of four contemporary LLMs: ChatGPT-5.2, Gemini 3, Copilot, and DeepSeek-V3.2 across major prosthodontic domains and question formats shows these models may support prosthodontic education and preliminary information retrieval, but cannot replace expert clinica...

Elif Yiğit İren, Hatice Betül Üçkuyu · 0 citations
Aug 2026

Comparative Evaluation of Large Language Models in Answering Patient Questions Following Periodontal and Peri-Implant Examination: An Expert-Based Study.

G Gemini demonstrated superior clinical precision and safety, whereas Claude provided more comprehensive and readable explanations, which support the integration of LLMs as pragmatic, high-ecological-validity complementary tools for patient education, while emphasizing the persistent necessity for professional clinical...

Ramazan Ağırağaç, Vedat Yüksekkaya · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.