Laura Wright

Author directory

2026

This study evaluates few-shot large language models (LLMs) on middle-school geoscience responses (N=86 responses × 11 indicators), separating presence agreement from extract agreement. Joint prompting and annotation-like rubric guidance plus examples yield the clearest gains; lengthy rubric rewriting and clause-level parsing do not reliably improve extract alignment.