Maher Jaoua
2026
Identifying Political Bias in Arabic News Articles
Saoussen Chaabane | Omar Trigui | Maher Jaoua
Proceedings of the 3rd Workshop on Natural Language Processing for Political Sciences (PoliticalNLP 2026)
Saoussen Chaabane | Omar Trigui | Maher Jaoua
Proceedings of the 3rd Workshop on Natural Language Processing for Political Sciences (PoliticalNLP 2026)
A comprehensive framework was developed to detect political bias in Arabic news articles, with a case study focusing on media reporting of the Palestinian issue. The methodology integrates MARBERT contextual embeddings with classical and deep learning classifiers, including SVM, Logistic Regression, Random Forest, and LSTM. The scalability of data processing was ensured through Apache Spark for potential real-time deployment. Experimental results showed that fine-tuned MARBERT embeddings combined with LSTM achieved the highest classification accuracy of 0.87, along with notable improvements in F1-scores across the pro, against, and neutral categories. These findings highlight the effectiveness of domain-specific fine-tuning of transformer models for political bias classification. The study also addressed class imbalance using SMOTE and class weighting strategies, and assessed feature robustness using multiple vectorization techniques.
2017
Machine Learning Approach to Evaluate MultiLingual Summaries
Samira Ellouze | Maher Jaoua | Lamia Hadrich Belguith
Proceedings of the MultiLing 2017 Workshop on Summarization and Summary Evaluation Across Source Types and Genres
Samira Ellouze | Maher Jaoua | Lamia Hadrich Belguith
Proceedings of the MultiLing 2017 Workshop on Summarization and Summary Evaluation Across Source Types and Genres
The present paper introduces a new MultiLing text summary evaluation method. This method relies on machine learning approach which operates by combining multiple features to build models that predict the human score (overall responsiveness) of a new summary. We have tried several single and “ensemble learning” classifiers to build the best model. We have experimented our method in summary level evaluation where we evaluate each text summary separately. The correlation between built models and human score is better than the correlation between baselines and manual score.
2013
An Evaluation Summary Method Based on a Combination of Content and Linguistic Metrics
Samira Ellouze | Maher Jaoua | Lamia Hadrich Belguith
Proceedings of the International Conference Recent Advances in Natural Language Processing RANLP 2013
Samira Ellouze | Maher Jaoua | Lamia Hadrich Belguith
Proceedings of the International Conference Recent Advances in Natural Language Processing RANLP 2013
An evaluation summary method based on combination of automatic and textual complexity metrics (Une méthode d’évaluation des résumés basée sur la combinaison de métriques automatiques et de complexité textuelle) [in French]
Samira Walha Ellouze | Maher Jaoua | Lamia Hadrich Belguith
Proceedings of TALN 2013 (Volume 2: Short Papers)
Samira Walha Ellouze | Maher Jaoua | Lamia Hadrich Belguith
Proceedings of TALN 2013 (Volume 2: Short Papers)
2008
Intégration d’une étape de pré-filtrage et d’une fonction multiobjectif en vue d’améliorer le système ExtraNews de résumé de documents multiples
Fatma Kallel Jaoua | Lamia Hadrich Belguith | Maher Jaoua | Abdelmajid Ben Hamadou
Actes de la 15ème conférence sur le Traitement Automatique des Langues Naturelles. Articles longs
Fatma Kallel Jaoua | Lamia Hadrich Belguith | Maher Jaoua | Abdelmajid Ben Hamadou
Actes de la 15ème conférence sur le Traitement Automatique des Langues Naturelles. Articles longs
Dans cet article, nous présentons les améliorations que nous avons apportées au système ExtraNews de résumé automatique de documents multiples. Ce système se base sur l’utilisation d’un algorithme génétique qui permet de combiner les phrases des documents sources pour former les extraits, qui seront croisés et mutés pour générer de nouveaux extraits. La multiplicité des critères de sélection d’extraits nous a inspiré une première amélioration qui consiste à utiliser une technique d’optimisation multi-objectif en vue d’évaluer ces extraits. La deuxième amélioration consiste à intégrer une étape de pré-filtrage de phrases qui a pour objectif la réduction du nombre des phrases des textes sources en entrée. Une évaluation des améliorations apportées à notre système est réalisée sur les corpus de DUC’04 et DUC’07.