Does ChatGPT provide safe and reliable patient information related to hip and knee arthroplasty?

BACKGROUND: Patients have been resorting to online content as their main source of medical knowledge, notably on orthopedic interventions. However, Web-based research findings present a reverse correlation between quality of content and popularity. We sought to evaluate whether ChatGPT could provide an alternative and safe source of medical information for arthroplasty patients. METHODS: We gave 5 commonly Googled questions related to hip and knee arthroplasty to board-certified arthroplasty surgeons, fellows, and orthopedic surgery residents, as well as to ChatGPT 4.0. We anonymized all answers, which were then analyzed by an independent board-certified arthroplasty surgeon. We scored the answers for accuracy of content (6-point Likert scale) and completeness (3-point Likert scale), then compared the performance of all groups. RESULTS: Among human responders, the mean accuracy grade was 75% (standard deviation [SD] 17%), with a mean completeness grade of 69% (SD 22%). The fellows represented the strongest subgroup, with 75% of their answers scored above 5/6 for accuracy (mean 82%, SD 16%). We found that all human-generated answers had a statistically significant correlation between accuracy and completeness. ChatGPT had a mean accuracy grade of 93% (SD 12%), with a mean completeness of 93% (SD 14%). CONCLUSION: ChatGPT appears to be a safe tool for patients to access general arthroplasty information online. It outperformed all human responders on both accuracy and completeness of answers. The strongest human responder group was the arthroplasty fellows. Further work is required to clarify the tool's performance against other easily accessible online patient information sources.

Authors

Institutions

Publication Details

Journal
Canadian Journal of Surgery
Published
2026-09-29
DOI
https://doi.org/10.1503/cjs.018325
Primary Topic
Artificial Intelligence in Healthcare and Education
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Does ChatGPT provide safe and reliable patient information related to hip and knee arthroplasty?

Joëlle Deschênes-Bilodeau, John Antoniou, Charles Desgagné, Peter Staunton
Canadian Journal of Surgery
Artificial Intelligence in Healthcare and Education
article

Does ChatGPT provide safe and reliable patient information related to hip and knee arthroplasty?

Joëlle Deschênes-Bilodeau, John Antoniou, Charles Desgagné, Peter Staunton
article en

Abstract

BACKGROUND: Patients have been resorting to online content as their main source of medical knowledge, notably on orthopedic interventions. However, Web-based research findings present a reverse correlation between quality of content and popularity. We sought to evaluate whether ChatGPT could provide an alternative and safe source of medical information for arthroplasty patients. METHODS: We gave 5 commonly Googled questions related to hip and knee arthroplasty to board-certified arthroplasty surgeons, fellows, and orthopedic surgery residents, as well as to ChatGPT 4.0. We anonymized all answers, which were then analyzed by an independent board-certified arthroplasty surgeon. We scored the answers for accuracy of content (6-point Likert scale) and completeness (3-point Likert scale), then compared the performance of all groups. RESULTS: Among human responders, the mean accuracy grade was 75% (standard deviation [SD] 17%), with a mean completeness grade of 69% (SD 22%). The fellows represented the strongest subgroup, with 75% of their answers scored above 5/6 for accuracy (mean 82%, SD 16%). We found that all human-generated answers had a statistically significant correlation between accuracy and completeness. ChatGPT had a mean accuracy grade of 93% (SD 12%), with a mean completeness of 93% (SD 14%). CONCLUSION: ChatGPT appears to be a safe tool for patients to access general arthroplasty information online. It outperformed all human responders on both accuracy and completeness of answers. The strongest human responder group was the arthroplasty fellows. Further work is required to clarify the tool's performance against other easily accessible online patient information sources.

Canadian Journal of SurgeryVol. 69(5)
Montreal General Hospital (CA), McGill University (CA)
Gender equality
Openalex Percentile: Top 16%
Artificial Intelligence in Healthcare and Education
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Does ChatGPT provide safe and reliable patient information related to hip and knee arthroplasty? — Joëlle Deschênes-Bilodeau, John Antoniou, et al. · Canadian Journal of Surgery (2026) | TGRS Research Map | TGRS