Automatic Extraction of Textual and Phonemic Complexity for French Cued Speech

Magali Norré, Brigitte Bigi, Núria Gala, Ludivine Javourey Drevet, Thomas François


Abstract
This article presents the results of an analysis of a written corpus with the view of automatically generating it in French Cued Speech (CS). CS is a communication system developed for people with hearing impairment to complement speech reading at the phonetic level using hands. This visual communication mode uses handshapes in different positions near the face in combination with the mouthshape (called ’cues’ or ’keys’) to make the phonemes of spoken language look different from each other. Despite many studies demonstrating its benefits, there are few resources available for learning and practicing it, especially in French. As part of a wider project aimed at creating an online learning platform with automatically generated videos using an augmented reality system displaying a virtual coding, we propose to identify, extract, and analyze 41 textual and phonemic features that might be more complex to (de)code in French CS. For the automatic extraction of complexity, several tools are used: FABRA for readability, SPPAS for phonetization and CS key generation. The results show some strong correlations between readability features, few between phonemic variables, and few between the two types. An initial model is proposed for selecting texts to be recorded for learning French CS.
Anthology ID:
2026.readi-1.5
Volume:
Proceedings of the Joint Workshop on Readability and Text Simplification (READIxTSAR) @ LREC 2026
Month:
May
Year:
2026
Address:
Palma, Mallorca (Spain)
Editors:
Matthew Shardlow, Thomas François, Raquel Amaro, Jorge Baptista, Rémi Cardon, Eugénio Ribeiro, Horacio Saggion, Regina Stodden, Amalia Todirascu, Rodrigo Wilkens
Venues:
READI | TSAR | WS
SIG:
Publisher:
ELRA Language Resources Association (ELRA)
Note:
Pages:
61–73
Language:
External URL:
https://lrec.elra.info/lrec2026-ws-readixtsar-05
DOI:
10.63317/2efggr7hkgst
Bibkey:
Cite (ACL):
Magali Norré, Brigitte Bigi, Núria Gala, Ludivine Javourey Drevet, and Thomas François. 2026. Automatic Extraction of Textual and Phonemic Complexity for French Cued Speech. In Proceedings of the Joint Workshop on Readability and Text Simplification (READIxTSAR) @ LREC 2026, pages 61–73, Palma, Mallorca (Spain). ELRA Language Resources Association (ELRA).
Cite (Informal):
Automatic Extraction of Textual and Phonemic Complexity for French Cued Speech (Norré et al., READI-TSAR 2026)
Copy Citation: