Open-access Dataset on Acceptability Ratings of Korean Clausal Constructions by Humans and GPT Models

Gyu-Ho Shin, Soo-Hwan Lee, Chanyoung Lee


Abstract
The present study introduces a new, open-access dataset on acceptability ratings of Korean clausal constructions at the morphosyntax–semantics interface (dative, passive, and negative polarity item). The dataset comprises (i) linguistically controlled sentence materials, (ii) ratings from targeted adult populations (individuals in their 20s), and (iii) parallel ratings from GPT variants (including ChatGPT). Alongside the release, we assess the alignment between GPT- and human-derived ratings to probe the extent to which GPT architectures can approximate patterns of human sentence comprehension.
Anthology ID:
2026.lrec-1.23
Volume:
Proceedings of the Fifteenth Language Resources and Evaluation Conference
Month:
May
Year:
2026
Address:
Palma de Mallorca, Spain
Editors:
Stelios Piperidis, Núria Bel, Henk van den Heuvel, Nancy Ide, Simon Krek, Antonio Toral
Venue:
LREC
SIG:
Publisher:
ELRA Language Resource Association
Note:
Pages:
344–356
Language:
External URL:
https://lrec.elra.info/lrec2026-main-023
DOI:
10.63317/2icd7h29b849
Bibkey:
Cite (ACL):
Gyu-Ho Shin, Soo-Hwan Lee, and Chanyoung Lee. 2026. Open-access Dataset on Acceptability Ratings of Korean Clausal Constructions by Humans and GPT Models. In Proceedings of the Fifteenth Language Resources and Evaluation Conference, pages 344–356, Palma de Mallorca, Spain. ELRA Language Resource Association.
Cite (Informal):
Open-access Dataset on Acceptability Ratings of Korean Clausal Constructions by Humans and GPT Models (Shin et al., LREC 2026)
Copy Citation: