AI paper index
Patient Facing AI Chatbots in Digital Orthodontic Education: A Comparative Evaluation of Understandability, Actionability, Quality, and Safety of Dietary Advice
One-line summary
An AI research paper on Patient Facing AI Chatbots in Digital Orthodontic Education: A Comparative Evaluation of Understandability, Actionability, Quality, and Safety of Dietary Advice.
Engineering notes
Engineering notes will be added by the aipentium editorial team.
Chinese explanation / 中文解读
中文解读待补充:本站会优先为大语言模型、生成式AI、ChatGPT相关技术、计算机视觉、深度学习等高价值论文补充中文说明。
Original abstract
Background/Objectives: Patient-facing artificial intelligence chatbots are increasingly used as informal digital health education tools. In orthodontics, eating and drinking advice may directly affect appliance integrity, oral hygiene, enamel demineralization, caries risk, and clear aligner use. This study compared the understandability, actionability, overall quality, and potential harmfulness of responses generated by free and paid versions of ChatGPT and Gemini to patient-oriented orthodontic dietary questions. Methods: This cross-sectional comparative observational study evaluated 160 responses generated from 40 Turkish patient-oriented orthodontic eating and drinking questions across five clinically relevant categories. Each question was submitted separately to ChatGPT free version, ChatGPT Plus, Gemini free version, and Gemini Pro on 1 May 2026, using newly opened independent chat sessions without prompt engineering, follow-up prompts, response regeneration, or manual editing. Responses were anonymized, randomly coded, and independently evaluated by two specialist dentists. PEMAT-P actionability was defined as the primary outcome. PEMAT-P understandability, Global Quality Score, and potentially harmful advice classification were secondary outcomes. Repeated-measures comparisons were performed using Friedman tests and Bonferroni-adjusted Wilcoxon signed-rank tests. Results: PEMAT-P understandability was high across all groups, with median scores of 100.0 in every group. Significant group differences were found for understandability, actionability, and Global Quality Score. ChatGPT Plus achieved the highest actionability score and Global Quality Score and produced no responses classified as potentially harmful. Potentially harmful responses were identified in ChatGPT free version, Gemini free version, and Gemini Pro. For the primary outcome, PEMAT-P actionability, the overall group difference was statistically significant with a small effect size (χ2 = 20.527, p < 0.001, Kendall’s W = 0.171), while the largest effect was observed for GQS with a moderate effect size (χ2 = 46.520, p < 0.001, Kendall’s W = 0.388). Conclusions: All chatbot groups generated highly understandable responses; however, actionability, overall quality, and safety varied across systems. ChatGPT Plus showed the strongest overall performance under the specific interface, subscription, language, and date conditions tested; however, this finding should be interpreted as a time-specific benchmark rather than evidence that paid chatbot systems are intrinsically safer or more clinically reliable. Structured evaluation, transparent reporting, digital health equity considerations, and professional oversight remain necessary before AI-generated orthodontic dietary advice can be integrated into routine patient education.
Links and sources
Need this topic turned into a technical roadmap?
aipentium can prepare a custom AI literature review, code map, dataset map, and B2B technology assessment.
Request B2B AI research
Comments