SimLex-999 for Modern Greek

Leonidas Mylonadis, Jelke Bloem


Abstract
Human judgements of word similarity have been a core benchmark for the intrinsic evaluation of word embedding models, and continue to be used for assessing the capabilities of large language models. While word similarity benchmarks have been collected for a range of languages, none existed for Greek. We develop a Modern Greek variant of the SimLex-999 word similarity dataset by gathering similarity judgements from 90 native speakers of Greek. We then use this as a benchmark for intrinsically evaluating several Greek language models.
Anthology ID:
2026.sigul-1.13
Volume:
Proceedings of the SIGUL 2026 Joint Workshop with ELE, EURALI, and DCLRL: Towards Inclusivity and Equality: Language Resources and Technologies for Under-Resourced and Endangered Languages
Month:
May
Year:
2026
Address:
Palma, Mallorca, Spain
Editors:
Atul Kr. Ojha, Sakriani Sakti, Claudia Soria, Maite Melero, John P. McCrae, Constantine Lignos, Chao-Hong Liu, German Rigau Claramunt, Georg Rehm
Venues:
SIGUL | EURALI | DCLRL | WS
SIG:
Publisher:
ELRA Language Resources Association (ELRA)
Note:
Pages:
119–125
Language:
External URL:
https://lrec.elra.info/lrec2026-ws-sigul-13
DOI:
10.63317/2ynm43iouxft
Bibkey:
Cite (ACL):
Leonidas Mylonadis and Jelke Bloem. 2026. SimLex-999 for Modern Greek. In Proceedings of the SIGUL 2026 Joint Workshop with ELE, EURALI, and DCLRL: Towards Inclusivity and Equality: Language Resources and Technologies for Under-Resourced and Endangered Languages, pages 119–125, Palma, Mallorca, Spain. ELRA Language Resources Association (ELRA).
Cite (Informal):
SimLex-999 for Modern Greek (Mylonadis & Bloem, SIGUL-EURALI-DCLRL 2026)
Copy Citation: