The performance of ChatGPT and other large language models on multiple-choice questions in biomedical disciplines: A meta-analysis.

Cheverko CM, Mavrych V, Bolgova O, Mohamed FRR, Westrick J, Juarez L, Rush E, Solka KA, Doubleday AF, Byram JN, Becker R, Gomez V, Ganeng BKA, Hoffman LA, Roach VA, Brown KM, DeVaul N, Garnett CN, Herriott HL, Lufler RS, Mussell JC, Balta JY, Pascoe MA, Middleton JW, Duffy S, Stephens GC, Wilson AB.

Open source

DOI
10.1002/ase.70262
Published
2026-05-19
Container
Anat Sci Educ
Publisher
Not recorded
Open access
yes

Credibility signals

limited evidence Score 45/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.1002/ase.70262,
  title = {The performance of ChatGPT and other large language models on multiple-choice questions in biomedical disciplines: A meta-analysis.},
  author = {Cheverko CM and  Mavrych V and  Bolgova O and  Mohamed FRR and  Westrick J and  Juarez L and  Rush E and  Solka KA and  Doubleday AF and  Byram JN and  Becker R and  Gomez V and  Ganeng BKA and  Hoffman LA and  Roach VA and  Brown KM and  DeVaul N and  Garnett CN and  Herriott HL and  Lufler RS and  Mussell JC and  Balta JY and  Pascoe MA and  Middleton JW and  Duffy S and  Stephens GC and  Wilson AB.},
  year = {2026},
  journal = {Anat Sci Educ},
  doi = {10.1002/ase.70262},
  url = {https://doi.org/10.1002/ase.70262}
}

RIS

TY  - JOUR
TI  - The performance of ChatGPT and other large language models on multiple-choice questions in biomedical disciplines: A meta-analysis.
AU  - Cheverko CM
AU  -  Mavrych V
AU  -  Bolgova O
AU  -  Mohamed FRR
AU  -  Westrick J
AU  -  Juarez L
AU  -  Rush E
AU  -  Solka KA
AU  -  Doubleday AF
AU  -  Byram JN
AU  -  Becker R
AU  -  Gomez V
AU  -  Ganeng BKA
AU  -  Hoffman LA
AU  -  Roach VA
AU  -  Brown KM
AU  -  DeVaul N
AU  -  Garnett CN
AU  -  Herriott HL
AU  -  Lufler RS
AU  -  Mussell JC
AU  -  Balta JY
AU  -  Pascoe MA
AU  -  Middleton JW
AU  -  Duffy S
AU  -  Stephens GC
AU  -  Wilson AB.
PY  - 2026
JO  - Anat Sci Educ
DO  - 10.1002/ase.70262
UR  - https://doi.org/10.1002/ase.70262
ER  - 

APA

CM, C., V, M., O, B., FRR, M., J, W., L, J., E, R., KA, S., AF, D., JN, B., R, B., V, G., BKA, G., LA, H., VA, R., KM, B., N, D., CN, G., HL, H., RS, L., JC, M., JY, B., MA, P., JW, M., S, D., GC, S., & AB., W. (2026). The performance of ChatGPT and other large language models on multiple-choice questions in biomedical disciplines: A meta-analysis.. Anat Sci Educ. https://doi.org/10.1002/ase.70262

Source records