Large language models as grading assistants in public health education: a method-comparison study of essay-style exam assessment
Large language models (LLMs) are increasingly being considered for assessment support in health professions education; however, evidence of their performance in essay-style examinations remains limited. In particular, little is known about the reproducibility and operational stability of LLM-based grading under dif...