NeCCo: Nepali Cultural Commonsense Benchmark for Large Language Model Evaluation

Sanket Shrestha, Raunak Regmi, Sadikshya Ghimire, Satyam Rana, Supriya Khadka


Abstract
Large language models perform strongly on standard evaluations, yet these benchmarks prioritize high-resource languages and culturally dominant knowledge, leaving culture-specific commonsense underexamined. In low-resource languages such as Nepali, everyday communication depends on culturally embedded cues, including kinship hierarchies, ritual practices, food systems, idioms, and honorific distinctions that literal translation often fails to capture. As a result, models that appear competent on global metrics can perform poorly in local contexts. To address this gap, we introduce NeCCo, a curated multiple-choice benchmark for culturally situated reasoning across five domains: kinship and social hierarchy; festivals, rituals, and geography; idioms, proverbs, and metaphors; commonsense and daily life; and gastronomy, agriculture, and nature. The dataset was created through structured authoring, cross-review, and normalization, and is released in Devanagari, English, and Romanized formats. We evaluate multiple state-of-the-art LLMs using standardized prompting and controlled decoding. Results show substantial variation: models perform better on globally documented knowledge such as geography, but struggle with relational and linguistically implicit tasks, including extended kinship reasoning and proverb interpretation. The most culturally dense categories expose brittleness and increased hallucination. These findings suggest that multilingual competence requires more than translation coverage and highlight the need for culturally grounded benchmarks and training signals.
Anthology ID:
2026.chipsal-1.15
Volume:
Proceedings of the Second workshop on Challenges in Processing South Asian Languages (CHiPSAL2026)
Month:
May
Year:
2026
Address:
Palma de Mallorca, Spain
Editors:
Kengatharaiyer Sarveswaran, Ashwini Vaidya
Venues:
CHiPSAL | WS
SIG:
Publisher:
ELRA Language Resources Association (ELRA)
Note:
Pages:
154–168
Language:
External URL:
https://lrec.elra.info/lrec2026-ws-chipsal-15
DOI:
10.63317/3zsmcs7pxtgi
Bibkey:
Cite (ACL):
Sanket Shrestha, Raunak Regmi, Sadikshya Ghimire, Satyam Rana, and Supriya Khadka. 2026. NeCCo: Nepali Cultural Commonsense Benchmark for Large Language Model Evaluation. In Proceedings of the Second workshop on Challenges in Processing South Asian Languages (CHiPSAL2026), pages 154–168, Palma de Mallorca, Spain. ELRA Language Resources Association (ELRA).
Cite (Informal):
NeCCo: Nepali Cultural Commonsense Benchmark for Large Language Model Evaluation (Shrestha et al., CHiPSAL 2026)
Copy Citation: