StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis.

Song X, Lee L, Xie K, Liu X, Deng X, Hong Y

Open source

DOI
10.1038/s41597-026-06731-4
Published
2026 Feb 6
Container
Scientific data
Publisher
Not recorded
Open access
yes

Credibility signals

uncertain Score 53/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.1038/s41597-026-06731-4,
  title = {StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis.},
  author = {Song X and Lee L and Xie K and Liu X and Deng X and Hong Y},
  year = {2026},
  journal = {Scientific data},
  doi = {10.1038/s41597-026-06731-4},
  url = {https://doi.org/10.1038/s41597-026-06731-4}
}

RIS

TY  - JOUR
TI  - StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis.
AU  - Song X
AU  - Lee L
AU  - Xie K
AU  - Liu X
AU  - Deng X
AU  - Hong Y
PY  - 2026
JO  - Scientific data
DO  - 10.1038/s41597-026-06731-4
UR  - https://doi.org/10.1038/s41597-026-06731-4
ER  - 

APA

X, S., L, L., K, X., X, L., X, D., & Y, H. (2026). StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis.. Scientific data. https://doi.org/10.1038/s41597-026-06731-4

Source records