Synapse
⌘+K
Synapse
PulseExploreClubsResearchersJournals
Instagram
HomeClubsExplore
October 27, 2025Northwestern Medical JournalOpen Access

Evaluation of ChatGPT-4.5 and DeepSeek-V3-R1 in answering patient-centered questions about orthognathic surgery: a comparative study across two languages

View Full Paper
Ask AI
Bookmark
Share

Authors

İSİpek Necla Güldiken SarıkayaEDEmrah Dilaver

Discussion

Loading...

Member takes

Overview

Comparative assessment reveals quality of responses from AI models in orthognathic surgery patient questions, highlighting reliability issues.

Key Points

  • Significant differences noted in quality across models answering orthognathic surgery questions, revealing accuracy gaps.
  • Statistical analyses showed notable results with non-parametric tests and inter-rater reliability assessments, including Cohen’s Kappa.
  • Assessment covered 25 patient-centered questions across two languages, offering insight into model performance disparities.
  • Artificial intelligence has potential but requires rigorous validation for delivering reliable patient information.

Cite This Study

Sarıkaya et al. (2025) studied this question.

synapsesocial.com/papers/68ff87d8c8c50a61f2bdcd0bhttps://doi.org/10.54307/2025.nwmj.220
View Full Paper
Ask AI
Bookmark
Share