ArticleScientific reports2026
Evaluation of large Language model performance on Persian rheumatology board exams: accuracy and clinical reasoning of GPT-4o vs. GPT-5.1.
Farzad Rafiei et al.PubMed ↗Full text ↗Publisher ↗
No numbers read from the abstract.
1 paper cites it
this papercites it