Reem Juhaysh


2026

Arabic handwritten manuscript recognition is challenging due to the cursive nature of the script, dot ambiguity, and document degradation. In this work, we propose an end-to-end OCR system based on a CNN–BiLSTM–CTC architecture. The model extracts visual features, captures sequential dependencies, and performs alignment-free training. Arabic-specific decoding and post-processing techniques are applied to reduce character and spacing errors. Experimental results show competitive performance in recognizing complex handwritten Arabic text.