Accessibility settings

Published on in Vol 28 (2026)

Preprints (earlier versions) of this paper are available at https://preprints.jmir.org/preprint/98184, first published .
Two data scientists analyze neural network code on screens, showing glowing brain and data visualizations.

The Reliability of Human Evaluation of Large Language Models in Health Care Settings: Scoping Review

The Reliability of Human Evaluation of Large Language Models in Health Care Settings: Scoping Review

Authors of this article:

Euijun Yang1 Author Orcid Image ;   Siyeon Ko1 Author Orcid Image ;   Hyekyung Woo1, 2 Author Orcid Image

Euijun Yang   1 , BS ;   Siyeon Ko   1 , MPH ;   Hyekyung Woo   1, 2 , PhD

1 Department of Health Administration, College of Nursing & Health, Kongju National University, Gongju, Chungcheongnam-do, Republic of Korea

2 Institute of Health and Environment, Kongju National University, Gongju, Chungcheongnam-do, Republic of Korea

Corresponding Author:

  • Hyekyung Woo, PhD
  • Department of Health Administration
  • College of Nursing & Health
  • Kongju National University
  • 56 Gongjudaehak-ro
  • Gongju, Chungcheongnam-do 32588
  • Republic of Korea
  • Phone: 82 41-850-0328
  • Email: hkwoo@kongju.ac.kr