Accessibility settings

Published on in Vol 28 (2026)

Preprints (earlier versions) of this paper are available at https://preprints.jmir.org/preprint/94755, first published .
Alternative text does not exist

Evaluating Large Language Models in Clinical Audiology (AUDIOLOGYBENCH): Benchmark Development and Validation Study

Evaluating Large Language Models in Clinical Audiology (AUDIOLOGYBENCH): Benchmark Development and Validation Study

Linkai Li   1, 2 * , MSc ;   Changgeng Mo   2 * , PhD ;   Haoshuai Zhou   2 , MSc ;   Hanlin Yu   3 , BASc ;   Congxi Lu   2 , MSc ;   Shangqiguo Wang   4 , PhD ;   Varsha M Athreya   5 , PhD ;   Matthew B Fitzgerald   5 , PhD ;   Shan X Wang   1 , PhD

1 Department of Electrical Engineering, Stanford University, Stanford, CA, United States

2 Orka Labs Inc, Shanghai, China

3 Department of Electrical and Computer Engineering, University of British Columbia, Vancouver, BC, Canada

4 Faculty of Education, University of Hong Kong, Hong Kong, Hong Kong, China (Hong Kong)

5 Department of Otolaryngology-Head and Neck Surgery, Stanford University, Stanford, CA, United States

*these authors contributed equally

Corresponding Author:

  • Shan X Wang, PhD
  • Department of Electrical Engineering, Stanford University
  • Address: McCullough Building, Room 351, 476 Lomita Mall
  • Stanford, CA
  • United States
  • Phone: 1 650-723-8671
  • Email: sxwang@stanford.edu