Published on in Vol 12 (2026)
Preprints (earlier versions) of this paper are
available at
https://preprints.jmir.org/preprint/72839, first published
.

Journals
- Zhang H, Qu L, Bai H, Chen Y, Ji R, Cheng Z, Yang C. Beyond accuracy: evaluating the reliability of large language models for medical assessment. Frontiers in Artificial Intelligence 2026;9 View
