Performance of large language models (GPT-5.2, Gemini 3 Pro, Claude Sonnet 4.6 and Grok 4.1) on the Fellowship of The Royal College of Surgeons Urology Part A examination.

Kafagi AR, Atassi N, Kafagi AH

Open source

DOI
10.1136/bmjopen-2025-108775
Published
2026 May 13
Container
BMJ open
Publisher
Not recorded
Open access
yes

Credibility signals

limited evidence Score 45/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.1136/bmjopen-2025-108775,
  title = {Performance of large language models (GPT-5.2, Gemini 3 Pro, Claude Sonnet 4.6 and Grok 4.1) on the Fellowship of The Royal College of Surgeons Urology Part A examination.},
  author = {Kafagi AR and Atassi N and Kafagi AH},
  year = {2026},
  journal = {BMJ open},
  doi = {10.1136/bmjopen-2025-108775},
  url = {https://doi.org/10.1136/bmjopen-2025-108775}
}

RIS

TY  - JOUR
TI  - Performance of large language models (GPT-5.2, Gemini 3 Pro, Claude Sonnet 4.6 and Grok 4.1) on the Fellowship of The Royal College of Surgeons Urology Part A examination.
AU  - Kafagi AR
AU  - Atassi N
AU  - Kafagi AH
PY  - 2026
JO  - BMJ open
DO  - 10.1136/bmjopen-2025-108775
UR  - https://doi.org/10.1136/bmjopen-2025-108775
ER  - 

APA

AR, K., N, A., & AH, K. (2026). Performance of large language models (GPT-5.2, Gemini 3 Pro, Claude Sonnet 4.6 and Grok 4.1) on the Fellowship of The Royal College of Surgeons Urology Part A examination.. BMJ open. https://doi.org/10.1136/bmjopen-2025-108775

Source records