Writer verification of partially damaged handwritten Arabic documents based on individual character shapes [PDF]
Author verification of handwritten text is required in several application domains and has drawn a lot of attention within the research community due to its importance.
Majid A. Khan +4 more
doaj +3 more sources
Revisiting K-Means and Topic Modeling, a Comparison Study to Cluster Arabic Documents
Clustering Arabic text documents is of high importance for many natural language technologies. This paper uses a combined method to cluster Arabic text documents. Mainly, we use generative models and clustering techniques. The study uses latent Dirichlet
M. Alhawarat, M. Hegazi
doaj +3 more sources
Plagiarism detection across languages: a comprehensive study of Arabic and English-to-Arabic long documents [PDF]
Plagiarism detection in Arabic texts remains a significant challenge due to the complex morphological structure, rich linguistic diversity, and scarcity of high-quality labeled datasets.
Ahmad Abdelaal +4 more
doaj +3 more sources
Examples of Arabic Documents in Ottoman Archives
The aim of this article is to inform researchers that there are many documents in the Ottoman Archives prepared in the Arabic language. Among them, the most influential collections are Waqf deeds. The number of Arabic Waqf deeds is more than we expected.
Mustafa Lütfi Bilge
doaj +1 more source
Computational Analysis of Printed Arabic Text Database for Natural Language Processing
A frequency dictionary of printed Arabic text is essential for natural language processing. It includes 1,251 XML files of Arabic documents collected from ten newspapers and magazines from different countries and created as the PATD database. A total of
Hassina Bouressace
doaj +1 more source
Multi-Agents Indexing System (MAIS) for Plagiarism Detection
With the extensive availability of technological systems all over the world and the increasing diffusion of information and documents by users, especially for the Arabic population, the development of semantic plagiarism detection systems has become ...
Samia Zouaoui, Khaled Rezeg
doaj +1 more source
A Superior Arabic Text Categorization Deep Model (SATCDM)
Categorizing Arabic text documents is considered an important research topic in the field of Natural Language Processing (NLP) and Machine Learning (ML).
M. Alhawarat, Ahmad O. Aseeri
doaj +1 more source
Inconsistency of translating medical abbreviations and acronyms into the Arabic language [PDF]
The paper identifies the challenges of translating variants of medical documents that include acronyms and abbreviations, which are standard in medicine, pharmacology, and healthcare contexts and documents, thus intending to examine the inconsistency of ...
Siddig Mohammed
doaj +1 more source
Segmentation of Ancient Arabic Documents [PDF]
This chapter addresses the problem of ancient Arabic document segmentation. As ancient documents have neither a real physical structure nor a logical one, the segmentation will be limited to textual areas or to line extraction in the areas. Although this type of segmentation appears quite simple, its implementation remains a challenging task.
Belaïd, Abdel, Ouwayed, Nazih
openaire +2 more sources
Segmentation-free Word Spotting for Handwritten Arabic Documents
In this paper we present an unsupervised segmentation-free method for spotting and searching query, especially, for images documents in handwritten Arabic, for this, Histograms of Oriented Gradients (HOGs) are used as the feature vectors to represent the
Ghizlane Khaissidi +5 more
doaj +1 more source

