Results 91 to 100 of about 2,520 (188)
The present paper deals with methods employed to create a lemmatized edition of all the witnesses of Roman de Horn, King Horn and Horn Childe and Maiden Rimnild.
Pierandrea Gottardi
doaj +1 more source
A Probabilistic Morphology Model for German Lemmatization [PDF]
Lemmatization is a central task in many NLP applications. Despite this importance, the number of (freely) available and easy to use tools for German is very limited. To fill this gap, we developed a simple lemmatizer that can be trained on any lemmatized corpus. For a full form word the tagger tries to find the sequence of morphemes that is most likely
openaire +2 more sources
Method of lemmatizer selections in multiplexing lemmatization
O A Sychev, N A Penskoy
openaire +1 more source
POS Tagging and Lemmatization of Historical Varieties of Languages. The Challenge of Old Italian
The paper discusses the challenges of POS tagging and lemmatization of historical varieties of Italian, and reports for both tasks the results of experiments carried out in a classical supervised domain adaptation scenario using the diachronic and ...
Manuel Favaro +2 more
doaj +1 more source
BabyLemmatizer : A Lemmatizer and POS-tagger for Akkadian
We present a hybrid lemmatizer and POS-tagger for Akkadian, the language of the ancient Assyrians and Babylonians, documented from 2350 BCE to 100 CE. In our approach the text is first POS-tagged and lemmatized with TurkuNLP trained with human-verified labels, and then post-corrected with dictionary-based methods to improve the lemmatization quality ...
Sahala Aleksi +3 more
openaire +2 more sources
Specyfika staropolszczyzny a anotacja gramatyczna. O lematyzacji tekstu staropolskiego
SPECIFIC FEATURES OF OLD POLISH LANGUAGE AND GRAMMATICAL ANNOTATION: LEMMATIZATION OF OLD POLISH TEXTS The article is dedicated to grammatical annotation of Old Polish texts. It discusses the problem with the lemmatization of a medieval text.
Magdalena Wismont +3 more
doaj +1 more source
This record contains a full paper presented at the 3rd Conference on Language Technologies (JT-2002), held in Ljubljana, Slovenia, in October 2002.
openaire +2 more sources
Analyzing Multilingual French and Russian Text using NLTK, spaCy, and Stanza
This lesson covers tokenization, part-of-speech tagging, and lemmatization, as well as automatic language detection, for non-English and multilingual text.
Ian Goodale
doaj +1 more source
Build Fast and Accurate Lemmatization for Arabic
In this paper we describe the complexity of building a lemmatizer for Arabic which has a rich and complex derivational morphology, and we discuss the need for a fast and accurate lammatization to enhance Arabic Information Retrieval (IR) results.
openaire +3 more sources
The plague of 1720 and migration in Martigues (France) in the 17th and 18th centuries. [PDF]
Darlu P, Séguy I.
europepmc +1 more source

