EduFairBench: reproducible evaluation of large language models for educational assessment.

Villegas-Ch W, Mera-Navarrete A, Zúñiga-Tello F, Gutiérrez R

Open source

DOI
10.3389/frai.2026.1933445
Published
2026
Container
Frontiers in artificial intelligence
Publisher
Not recorded
Open access
yes

Credibility signals

uncertain Score 53/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.3389/frai.2026.1933445,
  title = {EduFairBench: reproducible evaluation of large language models for educational assessment.},
  author = {Villegas-Ch W and Mera-Navarrete A and Zúñiga-Tello F and Gutiérrez R},
  year = {2026},
  journal = {Frontiers in artificial intelligence},
  doi = {10.3389/frai.2026.1933445},
  url = {https://doi.org/10.3389/frai.2026.1933445}
}

RIS

TY  - JOUR
TI  - EduFairBench: reproducible evaluation of large language models for educational assessment.
AU  - Villegas-Ch W
AU  - Mera-Navarrete A
AU  - Zúñiga-Tello F
AU  - Gutiérrez R
PY  - 2026
JO  - Frontiers in artificial intelligence
DO  - 10.3389/frai.2026.1933445
UR  - https://doi.org/10.3389/frai.2026.1933445
ER  - 

APA

W, V., A, M., F, Z., & R, G. (2026). EduFairBench: reproducible evaluation of large language models for educational assessment.. Frontiers in artificial intelligence. https://doi.org/10.3389/frai.2026.1933445

Source records