Benchmarking Multi-National Value Alignment for Large Language Models

Chengyi Ju; Weijie Shi; Chengzhong Liu; Jiaming Ji; Jipeng Zhang; Ruiyuan Zhang; Jiajie Xu; Yaodong Yang (杨耀东); Sirui Han; Yike Guo

doi:10.18653/v1/2025.findings-acl.1028

Benchmarking Multi-National Value Alignment for Large Language Models

Chengyi Ju, Weijie Shi, Chengzhong Liu, Jiaming Ji, Jipeng Zhang, Ruiyuan Zhang, Jiajie Xu, Yaodong Yang, Sirui Han, Yike Guo

Abstract

Do Large Language Models (LLMs) hold positions that conflict with your country’s values? Occasionally they do! However, existing works primarily focus on ethical reviews, failing to capture the diversity of national values, which encompass broader policy, legal, and moral considerations. Furthermore, current benchmarks that rely on spectrum tests using manually designed questionnaires are not easily scalable. To address these limitations, we introduce NaVAB, a comprehensive benchmark to evaluate the alignment of LLMs with the values of five major nations: China, the United States, the United Kingdom, France, and Germany. NaVAB implements a national value extraction pipeline to efficiently construct value assessment datasets. Specifically, we propose a modeling procedure with instruction tagging to process raw data sources, a screening process to filter value-related topics and a generation process with a Conflict Reduction mechanism to filter non-conflicting values. We conduct extensive experiments on various LLMs across countries, and the results provide insights into assisting in the identification of misaligned scenarios. Moreover, we demonstrate that NaVAB can be combined with alignment techniques to effectively reduce value concerns by aligning LLMs’ values with the target country.

Anthology ID:: 2025.findings-acl.1028
Volume:: Findings of the Association for Computational Linguistics: ACL 2025
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 20042–20058
Language:
URL:: https://aclanthology.org/2025.findings-acl.1028/
DOI:: 10.18653/v1/2025.findings-acl.1028
Bibkey:
Cite (ACL):: Chengyi Ju, Weijie Shi, Chengzhong Liu, Jiaming Ji, Jipeng Zhang, Ruiyuan Zhang, Jiajie Xu, Yaodong Yang, Sirui Han, and Yike Guo. 2025. Benchmarking Multi-National Value Alignment for Large Language Models. In Findings of the Association for Computational Linguistics: ACL 2025, pages 20042–20058, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: Benchmarking Multi-National Value Alignment for Large Language Models (Ju et al., Findings 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.findings-acl.1028.pdf

PDF Cite Search Fix data