Comparative Efficacy of ChatGPT and Gemini in Addressing Patient Queries on Gonarthrosis and Total Knee Arthroplasty: A Randomized Controlled Trial
Journal of Knee Surgery, cilt.39, sa.3, ss.123-126, 2026 (SCI-Expanded, Scopus)
- Yayın Türü: Makale / Tam Makale
- Cilt numarası: 39 Sayı: 3
- Basım Tarihi: 2026
- Doi Numarası: 10.1055/a-2693-0756
- Dergi Adı: Journal of Knee Surgery
- Derginin Tarandığı İndeksler: Science Citation Index Expanded (SCI-EXPANDED), Scopus, EMBASE, MEDLINE, Health Research Premium Collection (ProQuest)
- Sayfa Sayıları: ss.123-126
- Anahtar Kelimeler: artificial intelligence, total knee arthroplasty, patient education, ChatGPT, Gemini, gonarthrosis
- Sağlık Bilimleri Üniversitesi Adresli: Evet
Özet
The emergence of artificial intelligence (AI) in health care has created novel opportunities for enhancing patient education and alleviating anxiety. This study seeks to evaluate the effectiveness of two leading AI platforms, ChatGPT and Gemini, in delivering accurate and satisfactory responses to patients with gonarthrosis, considering total knee arthroplasty (TKA). A prospective, randomized controlled trial was conducted involving 100 patients diagnosed with gonarthrosis and indicated for TKA. Each patient posed five questions regarding the surgery and postoperative rehabilitation to both ChatGPT and Gemini. Responses were evaluated by two blinded orthopaedic specialists on a 10-point scale for accuracy and patient satisfaction. Patients additionally evaluated their satisfaction with each response using a 10-point scale. The main outcome measures consisted of the average accuracy scores assessed by specialists and the average satisfaction scores reported by patients. Statistical analysis revealed significant differences between ChatGPT and Gemini in both accuracy and patient satisfaction (p < 0.001). ChatGPT demonstrated better performance with a mean accuracy score of 8.7 ± 0.9 compared with Gemini's 7.2 ± 1.1. Patient satisfaction scores aligned with expert evaluations, with ChatGPT achieving a mean satisfaction score of 8.9 ± 0.8 versus Gemini's 7.5 ± 1.2. Notably, ChatGPT excelled in providing comprehensive explanations of surgical procedures (mean score: 9.2 ± 0.7) and postoperative care (9.1 ± 0.8), whereas Gemini performed better in offering concise summaries of recovery timelines (8.4 ± 0.9). This study demonstrates that ChatGPT offers more accurate and satisfactory responses to patient queries regarding gonarthrosis and TKA compared with Gemini. The findings suggest that AI platforms, particularly ChatGPT, can serve as valuable tools in augmenting patient education and potentially reducing preoperative anxiety. Future studies should investigate the incorporation of AI-assisted information delivery into clinical practice and its long-term effects on patient outcomes.