João Ricardo Silva
Papers on this page may belong to the following people: João Ricardo Silva, João Silva
2026
Progressing beyond Art Masterpieces or Touristic Clichés: how to assess your LLMs for cultural alignment?
António Branco | João Ricardo Silva | Nuno Marques | Luis M. S. Gomes | Ricardo Campos | Raquel Sequeira | Sara Nerea | Rodrigo Silva | Miguel Marques | Rodrigo Duarte | Artur Putyato | Diogo Folques | Tiago Valente
Proceedings of the Fourth Workshop on the Role of Resources in the Age of Large Language Models (RESOURCEFUL 2026)
António Branco | João Ricardo Silva | Nuno Marques | Luis M. S. Gomes | Ricardo Campos | Raquel Sequeira | Sara Nerea | Rodrigo Silva | Miguel Marques | Rodrigo Duarte | Artur Putyato | Diogo Folques | Tiago Valente
Proceedings of the Fourth Workshop on the Role of Resources in the Age of Large Language Models (RESOURCEFUL 2026)
Although the cultural (mis)alignment of Large Language Models (LLMs) has attracted increasing attention - often framed in terms of cultural bias - until recently there has been limited work on the design and development of datasets for cultural assessment. Here, we review existing approaches to such datasets and identify their main limitations. To address these issues, we propose design guidelines for annotators and report on the construction of a dataset built according to these principles. We further present a series of contrastive experiments conducted with this dataset. The results demonstrate that our design yields test sets with greater discriminative power, effectively distinguishing between models specialized for a given culture and those that are not, ceteris paribus.
CLARIN-PT-LDB: An Open LLM Leaderboard for Portuguese to assess Language, Culture and Civility
João Ricardo Silva | Luís Gomes | António Branco
Proceedings of the 17th International Conference on Computational Processing of Portuguese (PROPOR 2026) - Vol. 1
João Ricardo Silva | Luís Gomes | António Branco
Proceedings of the 17th International Conference on Computational Processing of Portuguese (PROPOR 2026) - Vol. 1
This paper reports on the development of a leaderboard of Open Large Language Models (LLM) for European Portuguese (PT-PT), and on its associated benchmarks.This leaderboard comes as a way to address a gap in the evaluation of LLM for European Portuguese, which so far had no leaderboard dedicated to this variant of the language.The paper also reports on novel benchmarks, including some that address aspects of performance that so far have not been available in benchmarks for European Portuguese, namely model safeguards and alignment to Portuguese culture.The leaderboard is available at https://huggingface.co/spaces/PORTULAN/portuguese-llm-leaderboard.