Assessment of the reliability and usability of ChatGPT in response to spinal cord injury questions
Journal of Spinal Cord Medicine, cilt.48, sa.5, ss.852-857, 2025 (SCI-Expanded, Scopus)
- Yayın Türü: Makale / Tam Makale
- Cilt numarası: 48 Sayı: 5
- Basım Tarihi: 2025
- Doi Numarası: 10.1080/10790268.2024.2361551
- Dergi Adı: Journal of Spinal Cord Medicine
- Derginin Tarandığı İndeksler: Science Citation Index Expanded (SCI-EXPANDED), Scopus, CINAHL, EMBASE, MEDLINE
- Sayfa Sayıları: ss.852-857
- Anahtar Kelimeler: ChatGPT, Spinal cord injury, Reliability, Artificial intelligence
- Sağlık Bilimleri Üniversitesi Adresli: Evet
Özet
Objective: The use of artificial intelligence chatbots to obtain information about patients’ diseases is increasing. This study aimed to determine the reliability and usability of ChatGPT for spinal cord injury-related questions. Methods: Three raters simultaneously evaluated a total of 47 questions on a 7-point Likert scale for reliability and usability, based on the three most frequently searched keywords in Google Trends (‘general information’, ‘complications’ and ‘treatment’). Results: Inter-rater Cronbach α scores indicated substantial agreement for both reliability and usability scores (α between 0.558 and 0.839, and α between 0.373 and 0.772, respectively). The highest mean reliability score was for ‘complications’ (mean 5.38). The lowest average was for the ‘general information’ section (mean 4.20). The ‘treatment’ had the highest mean scores for the usability (mean 5.87) and the lowest mean value was recorded in the ‘general information’ section (mean 4.80). Conclusion: The answers given by ChatGPT to questions related to spinal cord injury were reliable and useful. Nevertheless, it should be kept in mind that ChatGPT may provide incorrect or incomplete information, especially in the ‘general information’ section, which may mislead patients and their relatives.