Measuring the quality of AI-generated clinical notes: A systematic review and experimental benchmark of evaluation methods.

Dahlberg A, Käenniemi T, Winther-Jensen T, Tapiola O, Luisto R, Puranen T, Gordon M, Sanmark E, Vartiainen V.

Open source

DOI
10.1016/j.artmed.2026.103421
Published
2026-04-07
Container
Artif Intell Med
Publisher
Not recorded
Open access
no

Credibility signals

limited evidence Score 43/100 under policy 1.0.0. This is a metadata assessment, not a judgment of the paper's conclusions.

Show all credibility signals

Cite this work

BibTeX

@article{allodium:10.1016/j.artmed.2026.103421,
  title = {Measuring the quality of AI-generated clinical notes: A systematic review and experimental benchmark of evaluation methods.},
  author = {Dahlberg A and  Käenniemi T and  Winther-Jensen T and  Tapiola O and  Luisto R and  Puranen T and  Gordon M and  Sanmark E and  Vartiainen V.},
  year = {2026},
  journal = {Artif Intell Med},
  doi = {10.1016/j.artmed.2026.103421},
  url = {https://doi.org/10.1016/j.artmed.2026.103421}
}

RIS

TY  - JOUR
TI  - Measuring the quality of AI-generated clinical notes: A systematic review and experimental benchmark of evaluation methods.
AU  - Dahlberg A
AU  -  Käenniemi T
AU  -  Winther-Jensen T
AU  -  Tapiola O
AU  -  Luisto R
AU  -  Puranen T
AU  -  Gordon M
AU  -  Sanmark E
AU  -  Vartiainen V.
PY  - 2026
JO  - Artif Intell Med
DO  - 10.1016/j.artmed.2026.103421
UR  - https://doi.org/10.1016/j.artmed.2026.103421
ER  - 

APA

A, D., T, K., T, W., O, T., R, L., T, P., M, G., E, S., & V., V. (2026). Measuring the quality of AI-generated clinical notes: A systematic review and experimental benchmark of evaluation methods.. Artif Intell Med. https://doi.org/10.1016/j.artmed.2026.103421

Source records