Results 11 to 20 of about 226,334 (282)

Hybrid artificial intelligence architectures for automatic text correction in the Kazakh language [PDF]

open access: yesFrontiers in Artificial Intelligence
The Kazakh language, as an agglutinative and morphologically rich language, presents significant challenges for the development of natural language processing (NLP) tools.
Laura Baitenova   +4 more
doaj   +2 more sources

Effective Spell Checking Methods Using Clustering Algorithms [PDF]

open access: yes, 2013
This paper presents a novel approach to spell checking using dictionary clustering. The main goal is to reduce the number of times distances have to be calculated when finding target words for misspellings.
Cordeiro De Amorim, Renato   +1 more
core   +5 more sources

Text segmentation for Chinese spell checking [PDF]

open access: yesJournal of the American Society for Information Science, 1999
Chinese spell checking is different from its counterparts for Western languages because Chinese words in texts are not separated by spaces. Chinese spell checking in this article refers to how to identify the misuse of characters in text composition. In other words, it is error correction at the word level rather than at the character level.
Kin-Hong Lee   +2 more
openaire   +2 more sources

Autocomplete and Spell Checking Levenshtein Distance Algorithm To Getting Text Suggest Error Data Searching In Library

open access: yesScientific Journal of Informatics, 2018
Nowadays internet technology provide more convenience for searching information on a daily. Users are allowed to find and publish their resources on the internet using search engine.
Muhamad Maulana Yulianto   +2 more
doaj   +1 more source

Correcting Diacritics and Typos with a ByT5 Transformer Model

open access: yesApplied Sciences, 2022
Due to the fast pace of life and online communications and the prevalence of English and the QWERTY keyboard, people tend to forgo using diacritics, make typographical errors (typos) when typing in other languages.
Lukas Stankevičius   +4 more
doaj   +1 more source

Defining Semantically Close Words of Kazakh Language with Distributed System Apache Spark

open access: yesBig Data and Cognitive Computing, 2023
This work focuses on determining semantically close words and using semantic similarity in general in order to improve performance in information retrieval tasks.
Dauren Ayazbayev   +3 more
doaj   +1 more source

Spelling-checking for highly inflective languages [PDF]

open access: yesProceedings of the 13th conference on Computational linguistics -, 1990
Spelling-checkers have become an integral part of most text processing software. From different reasons among which the speed of processing prevails they are usually based on dictionaries of word forms instead of words. This approach is sufficient for languages with little inflection such as English, but fails for highly inflective languages such as ...
Jan Hajic 0001, Janus Drózd
openaire   +2 more sources

A UMLS-based spell checker for natural language processing in vaccine safety

open access: yesBMC Medical Informatics and Decision Making, 2007
Background The Institute of Medicine has identified patient safety as a key goal for health care in the United States. Detecting vaccine adverse events is an important public health activity that contributes to patient safety.
Liu Fang   +8 more
doaj   +1 more source

PRo-Pat: Probabilistic Root–Pattern Bi-gram data language model for Arabic based morphological analysis and distribution

open access: yesData in Brief, 2023
Based on 29,192,662 html files obtained from the ClueWeb a bi-gram data language model for Arabic is constructed. The created dataset is considering standard types of bi-gram analysis, however with focus on the root11 An Arabic root depict the basic ...
Bassam Haddad   +3 more
doaj   +1 more source

Improving Document Retrieval with Spelling Correction for Weak and Fabricated Indonesian-Translated Hadith

open access: yesJurnal RESTI (Rekayasa Sistem dan Teknologi Informasi), 2020
Hadith has several levels of authenticity, among which are weak (dhaif), and fabricated (maudhu) hadith that may not originate from the prophet Muhammad PBUH, and thus should not be considered in concluding an Islamic law (sharia).
muhammad zaky ramadhan   +1 more
doaj   +1 more source

Home - About - Disclaimer - Privacy