The Swedish Benchmark of Linguistic Minimal Pairs

Johan Sjons, Fredrik Heinat, Murathan Kurfali


Abstract
We introduce the Swedish Benchmark of Linguistic Minimal Pairs, a dataset for evaluating syntactic performance in language models. It includes 2,500 minimal pairs organized into 25 syntactic phenomena, with 100 pairs per phenomenon. Each pair contrasts a well-formed and an ill-formed sentence that differ minimally. For each phenomenon, we manually constructed ten pairs from scratch. We semi-automatically generated the remaining 90 pairs and manually adjusted them. A random sample was assessed by 40 participants, who selected the well-formed sentence in 98.05% of cases. We evaluate eleven state-of-the-art models. Results generally show that models handle local agreement well but struggle with certain long-distance dependencies and word order phenomena. Model size seems to matter less than the training domain. Prompt-based evaluation generally lowers performance. We show that model performance is stable across handcrafted and generated subsets and across sample sizes, suggesting that 100 pairs per phenomenon suffice for reliable evaluation. Future work will expand the number of phenomena.
Anthology ID:
2026.lrec-1.540
Volume:
Proceedings of the Fifteenth Language Resources and Evaluation Conference
Month:
May
Year:
2026
Address:
Palma de Mallorca, Spain
Editors:
Stelios Piperidis, Núria Bel, Henk van den Heuvel, Nancy Ide, Simon Krek, Antonio Toral
Venue:
LREC
SIG:
Publisher:
ELRA Language Resource Association
Note:
Pages:
6783–6794
Language:
External URL:
https://lrec.elra.info/lrec2026-main-540
DOI:
10.63317/33cfy28hybv5
Bibkey:
Cite (ACL):
Johan Sjons, Fredrik Heinat, and Murathan Kurfali. 2026. The Swedish Benchmark of Linguistic Minimal Pairs. In Proceedings of the Fifteenth Language Resources and Evaluation Conference, pages 6783–6794, Palma de Mallorca, Spain. ELRA Language Resource Association.
Cite (Informal):
The Swedish Benchmark of Linguistic Minimal Pairs (Sjons et al., LREC 2026)
Copy Citation: