Houda Elmimouni
2026
From Cairo to Cape Town: How African Twitter Shapes the Global Palestine-Israel Narrative
Mahmoud Fawzi | Houda Elmimouni | Walid Magdy
Proceedings of the 3rd Workshop on Natural Language Processing for Political Sciences (PoliticalNLP 2026)
Mahmoud Fawzi | Houda Elmimouni | Walid Magdy
Proceedings of the 3rd Workshop on Natural Language Processing for Political Sciences (PoliticalNLP 2026)
African Twitter users are active shapers of the Palestine-Israel conversation but their contribution remains relatively understudied. Using 132.5K geo-located tweets from 2020 to 2023 and 451-term list of keywords in 33 languages, we identify three patterns in this context: (1) broad participation (Egypt supplies 43% of posts, yet Nigeria, South Africa, Kenya and Ghana contribute more than a third); (2) multilingual predominantly pro-Palestine amplification across Arabic, English, French, Swahili and other tongues, with 8% of tweets left “undetermined” by Twitter’s language detector; and (3) a humanitarian framing that centers civilian harm through hashtags such as #GazaUnderAttack and #PalestenianLivesMatter. We outline design implications for language-agnostic interfaces, low-friction source verification and cross-movement recommendation tools that foreground African epistemologies in global civic-tech systems.
The NakbaEcho Dataset: From Oral Testimonies to a Transcribed Arabic History Corpus
Batool Najeh Balah | Mahmoud Fawzi | Houda Elmimouni | Walid Magdy
Proceedings of the 2nd International Workshop on Nakba Narratives as Language Resources @ LREC 2026
Batool Najeh Balah | Mahmoud Fawzi | Houda Elmimouni | Walid Magdy
Proceedings of the 2nd International Workshop on Nakba Narratives as Language Resources @ LREC 2026
We present NakbaEcho, a dataset derived from Palestinian testimonies about the 1948 Nakba. The resource is constructed from transcribing over 2,180 hours of recorded interviews gathered through the Palestine Remembered Oral History index and linked to multiple repositories, including the Palestinian Oral History Archive (POHA) and YouTube-hosted interviews. We harmonize interview-level metadata and generate timestamp-aligned transcripts from the original Arabic recordings using an automatic transcription pipeline configured for Palestinian Arabic. The dataset includes speaker-labeled segments and auxiliary annotations designed to support downstream research in Arabic speech processing, natural language processing, digital humanities, and oral-history analysis. NakbaEcho contributes a structured computational resource for studying Palestinian oral testimony while expanding the availability of dialectal Arabic materials for speech, text, and social research.