Jorge Rico
2026
Privacy VITA: a new multilingual and multimodal annotated video corpus to evaluate anonymization systems
Jorge Rico | Sofia Contreras | Enrique Manjavacas Arevalo | Maria Viana | Ruben Pérez-Ramón | Maria Luisa Izquierdo | María José Vilella | Jaime Corton | Silvia Rodriguez | Fernando Espinza | Rafael Ginard | César Pérez | Jesús Arias | Richard Cook | Pablo Regodón | Manuel Moyano | Ricardo Heredia | Lara De Santos | Pierre Plaza | Jacqueline González
Proceedings of the Second International Conference on Natural Language Processing and Artificial Intelligence for Cyber Security
Jorge Rico | Sofia Contreras | Enrique Manjavacas Arevalo | Maria Viana | Ruben Pérez-Ramón | Maria Luisa Izquierdo | María José Vilella | Jaime Corton | Silvia Rodriguez | Fernando Espinza | Rafael Ginard | César Pérez | Jesús Arias | Richard Cook | Pablo Regodón | Manuel Moyano | Ricardo Heredia | Lara De Santos | Pierre Plaza | Jacqueline González
Proceedings of the Second International Conference on Natural Language Processing and Artificial Intelligence for Cyber Security
In this paper, we introduce a new multimodal and multilingual public resource for evaluating anonymization models and solutions, comprising a collection of 516 annotated videos. Unlike previously available resources, the Privacy VITA corpus is both multilingual and multimodal, covering Video, Image, Text and Audio modalities. The development of this dataset is motivated by the growing need to anonymize private data in an increasingly multimedia-driven society. Furthermore, emerging regulations pose significant challenges for organizations and companies that manage sensitive information while ensuring legal compliance. This paper details the processes of collection, preprocessing, annotation, and curation of the corpus, as well as its overall scope. We hope this resource will contribute to the advancement of multimodal anonymization research.
Design and Methodological Architecture of a Multilingual Corpus of Interpreter-mediated Public Service Telephone Interactions
Raquel Lazaro Gutierrez | Daniel López Padilla | Jorge Rico | María José Vilella Sánchez | Fernando Manuel Espinoza-Cuadros
Proceedings of Shaping Multilingual, Multimodal AI for the Social Sciences and Humanities (LLMs4SSH) @ LREC 2026
Raquel Lazaro Gutierrez | Daniel López Padilla | Jorge Rico | María José Vilella Sánchez | Fernando Manuel Espinoza-Cuadros
Proceedings of Shaping Multilingual, Multimodal AI for the Social Sciences and Humanities (LLMs4SSH) @ LREC 2026
Multimodality in Social Sciences and Humanities (SSH) research is often associated with the integration of text and visual data. However, interpreter-mediated telephone interaction presents a different configuration of complexity, where acoustic, temporal, discursive, and pragmatic dimensions converge. This paper presents the design and methodological architecture of PRAGMACOR(Corpus Pragmatics and Telephone Interpreting: Analysis of Face-Threatening Acts, Ref. PID2021-127196NA-I00), a multilingual corpus of interpreter-mediated public service telephone interactions (Chinese–Spanish, English–Spanish, French–Spanish, German–Spanish), as a case study in multimodal and plurilingual SSH infrastructure. The corpus integrates aligned audio recordings, orthographic transcriptions enriched with speech phenomena, temporal segmentation into speech acts, and multilayer pragmatic annotation of Face-Threatening Acts (FTAs), validated through a structured double-annotation and expert review process. Beyond textual data, the infrastructure captures prosodic overlap, turn-taking dynamics, and pragmatic mediation, enabling the study of cross-linguistic transfer and relational negotiation in asymmetrical institutional contexts. Datasets such as PRAGMACOR have proved essential to train LLMs for speech to speech translation (Sakai et al., 2024). Attention is given to the ethical and technical design of the corpus, including local automatic transcription, systematic removal of personal identifiable information, and irreversible voice anonymization through spectral and temporal signal transformation. These procedures ensure both research usability and compliance with responsible data governance principles. By conceptualising interpreter-mediated interaction as an acoustic-discursive multimodal object and plurilingual pragmatic process, this paper argues that PRAGMACOR provides a replicable model for the development of SSH-oriented infrastructures capable of supporting advanced research in multilingual communication, discourse analysis, and future evaluation of language technologies.
Search
Fix author
Co-authors
- Jesús Arias 1
- Sofia Contreras 1
- Richard Cook 1
- Jaime Corton 1
- Lara De Santos 1
- Fernando Manuel Espinoza-Cuadros 1
- Fernando Espinza 1
- Rafael Ginard 1
- Jacqueline González 1
- Raquel Lázaro Gutiérrez 1
- Ricardo Heredia 1
- Maria Luisa Izquierdo 1
- Daniel López Padilla 1
- Enrique Manjavacas 1
- Manuel Moyano 1
- Pierre Plaza 1
- César Pérez 1
- Ruben Pérez-Ramón 1
- Pablo Regodón 1
- Silvia Rodriguez 1
- Maria Viana 1
- María José Vilella 1
- María José Vilella Sánchez 1