Transfer Learning from ImageNet-Pretrained Vision Models for Arabic Text Classification Using Glyph-Based Image Representations

Authors

DOI:

https://doi.org/10.54361/ajmas.2584134

Keywords:

Arabic Text Classification, Transfer Learning, Glyph-Based Representation, EfficientNet-B3, Xception

Abstract

Arabic text classification is challenging because of rich morphology, orthographic variation, and complex tokenization. This study investigates an alternative approach that represents Arabic documents as rendered images and classifies them using ImageNet-pretrained vision models. We developed a rendering pipeline to preserve contextual letter shaping and right-to-left layout; we also divided long documents into image tiles and classified them by averaging tile-level probabilities. EfficientNet-B3 and Xception were fine-tuned using the same two-stage training protocol without automated hyperparameter optimization. Experiments were conducted on 6,000 balanced Modern Standard Arabic news articles across six categories. EfficientNet-B3 achieved 98.77% test accuracy, compared with 98.34% for Xception, while requiring approximately half the parameters and one-fifth of the floating-point operations. The approach eliminates tokenization, vocabulary construction, and manual feature engineering while achieving competitive classification performance. These results demonstrate the potential of rendered Arabic text as a compact and effective representation for visual transfer learning in Arabic text classification.

Downloads

Published

2025-12-30

How to Cite

1.
Hosam Alzawam, Tarik Idbeaa. Transfer Learning from ImageNet-Pretrained Vision Models for Arabic Text Classification Using Glyph-Based Image Representations. Alq J Med App Sci [Internet]. 2025 Dec. 30 [cited 2026 Oct. 7];:3010-2. Available from: https://journal.utripoli.edu.ly/index.php/Alqalam/article/view/1988

Issue

Section

Articles