Joshua A McGrane
Author directory2026
How Much Training Data Is Enough? Fine-Tuning LLMs for Short-Answer Scoring
Joshua A McGrane | David Torres Irribarra
Proceedings of the Artificial Intelligence in Measurement and Education Conference (AIME-Con): Works in Progress
Joshua A McGrane | David Torres Irribarra
Proceedings of the Artificial Intelligence in Measurement and Education Conference (AIME-Con): Works in Progress
Fine-tuned models approached state-of-the-art agreement using small labelled sets. Open-weight models met operational criteria on seven of ten items, with a median of 50 responses per score point among passing items. A single marking exercise may supply enough data for automated short-answer scoring.