TryggLLM: A Benchmark for Evaluating LLM Safety in Norwegian

Samia Touileb, Truls Pedersen, Isabell Stinessen Haugen


Abstract
We introduce TryggLLM, the first safety benchmark dataset for Norwegian. The dataset is intended for benchmarking different types of safety issues that can occur when using Norwegian generative language models. We have manually translated two English benchmark datasets, while modifying the content to be aligned with the Norwegian context. The benchmark dataset is composed of two sub-parts: i) prompts annotated by four native speakers, in both the written variants of Norwegian Bokmål (BM) and Nynorsk (NN), such that each native speaker wrote in their preferred variants (two BM and two NN); ii) prompts and target responses, where each of them has a BM and a NN version. We provide detailed descriptions of the data creation process. We also present a thorough manual evaluation of benchmarking existing open Norwegian LLMs using TryggLLM. Our results show that between 18% and 48% of the generated responses are unsafe, across all tested models.
Anthology ID:
2026.lrec-1.323
Volume:
Proceedings of the Fifteenth Language Resources and Evaluation Conference
Month:
May
Year:
2026
Address:
Palma de Mallorca, Spain
Editors:
Stelios Piperidis, Núria Bel, Henk van den Heuvel, Nancy Ide, Simon Krek, Antonio Toral
Venue:
LREC
SIG:
Publisher:
ELRA Language Resource Association
Note:
Pages:
4093–4102
Language:
External URL:
https://lrec.elra.info/lrec2026-main-323
DOI:
10.63317/2rhfg2a92wim
Bibkey:
Cite (ACL):
Samia Touileb, Truls Pedersen, and Isabell Stinessen Haugen. 2026. TryggLLM: A Benchmark for Evaluating LLM Safety in Norwegian. In Proceedings of the Fifteenth Language Resources and Evaluation Conference, pages 4093–4102, Palma de Mallorca, Spain. ELRA Language Resource Association.
Cite (Informal):
TryggLLM: A Benchmark for Evaluating LLM Safety in Norwegian (Touileb et al., LREC 2026)
Copy Citation: