article
This paper presents a bilingual pipeline for Arabic text summarization, addressing the unique academic and technical content’s linguistic challenges. The proposed pipeline solves the main problems associated with Arabic-specific models (e.g. AraBART, AraT5, and mBERT2mBERT) for summarization. Comparative evaluations using the XL-Sum dataset demonstrate that the proposed bilingual pipeline outperformed all tested models with a ROUGE-2 F1 score of 39.44% versus AraBART’s 33%, the most known one for Arabic text summarization. This research emphasizes the huge performance gap between English-language transformers and their Arabic counterparts, highlighting the potential of bilingual approaches to bridge this gap. By leveraging both English and Arabic models, the proposed methods significantly create a concise and coherent summarization that can be used in many other fields, particularly in academia.
This page summarises published work. The authoritative version sits with the publisher.
DOI: 10.1109/src65875.2025.11263719
Is something wrong with this record? Report it or request removal.
Discussion
Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.
No discussion yet. Open the first thread.
New to MARATTO™? Create a free account.