Open access
Jul 2026
Evaluating large language models for rubric-based essay grading in an undergraduate biology course
It is suggested that LLM grading outputs vary meaningfully across models, prompting strategies, and rubric components, and this context, LLMs may be best understood as tools that can support specific aspects of structured grading rather than as interchangeable evaluators.
M. Naidu, Nikolas S. Montaquila, Jessica P Roa et al.
· Journal of Microbiology & Bi... · 0 citations