Benchmarking large language models on persian surgical subspecialty board examinations: a comparative study of ChatGPT-4o, ChatGPT-5, and Gemini 2.5 Flash

S Shahab Sheikhalishahi F Farzad Rafiei S Seyed Masoud Hosseini A Alireza Haddadi S Saina Sadeghipour

Article Details

Volume / Issue Vol. 16, Issue 1
Published May 08, 2026
ISSN 2045-2322
Publisher Nature Portfolio

Journal Info

Scientific Reports

Nature Portfolio

ISSN: 2045-2322 Open Access Life Sciences

Authors (5)

S

Shahab Sheikhalishahi

F

Farzad Rafiei

S

Seyed Masoud Hosseini

A

Alireza Haddadi

S

Saina Sadeghipour