1
Return

Four general-purpose large language models (ChatGPT-5, Claude 4, Grok 4 and Gemini 2.5) show comparable performance in specialised total knee arthroplasty clinical questions

delete2026-08-12
delete0
delete
OA
AI
O
Oriol Pujol
R
Robert Ferrer
A
Alex Coelho
F
Felix C. Oettl
B
Bálint Zsidai
J
Joan Leal-Blanquet
M
Michael T. Hirschmann
K
Kristian Samuelsson *
DOI:10.1002/ksa.70544delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
To evaluate and compare the performance of four general-purpose large language models (LLMs) (ChatGPT-5, Claude 4, Grok 4 and Gemini 2.5) in answering specialised clinical questions related to total knee arthroplasty (TKA) derived from the World Expert Meeting in Arthroplasty (WEMA).
Keywords:
artificial intelligence
large language model
total knee arthroplasty
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Knee Surgery Sports Traumatology Arthroscopy cover
Knee Surgery Sports Traumatology Arthroscopy
IF:
5
Papers:
9.0K
Citations:
2.5W

Organization

K
Kantonsspital Baselland
Scholars:
814
Papers: 663
Citations: 416
I
imove traumatology
Scholars:
2
Papers: 1
Citations: 0
S
sahlgrenska sports medicine center
Scholars:
21
Papers: 7
Citations: 1
U
university of zurich
Scholars:
4.9W
Papers: 3.9W
Citations: 65
F
fundacio hospitalaries hospital
Scholars:
2
Papers: 1
Citations: 0
V
Vall d'Hebron University Hospital
Scholars:
91
Papers: 34
Citations: 0
Cited Papers

Cited Papers

Citing Papers

Citing Papers