Results 11 to 20 of about 160 (101)
TEICORPO: A Conversion Tool for Spoken Language Transcription with a Pivot File in TEI
CORLI is a consortium of Huma-Num, the French national infrastructure dedicated to the technical support and promotion of digital humanities. The goal of CORLI is to promote and provide tools and information for good and efficient research practices in ...
Christophe Parisse +2 more
doaj +1 more source
Neural POS tagging of shahmukhi by using contextualized word representations
Part of Speech (POS) tagging has a preliminary role in building natural language processing applications. This paper presents the development and evaluation of the first POS tagged corpus along with a Bi-directional long-short memory (BiLSTM) network ...
Amina Tehseen +4 more
doaj +1 more source
En este artículo ofrecemos una formalización de reglas de pluralización en castellano para ser utilizada concretamente en el procesamiento de términos especializados, ya que con frecuencia estos no se encuentran registrados en los diccionarios de lengua ...
Rogelio Nazar, Amparo Galdames
doaj +1 more source
Noun phrase complexity in Ghanaian English
Abstract This study compares the complexity of the noun phrase (NP) in Ghanaian English in a real‐time perspective. Based on the Historical Corpus English in Ghana (1966–1975) and the Ghanaian component of the International Corpus of English (mainly 2000s), representing the early and late stages of structural nativisation in the dynamic model, NP ...
Thorsten Brato
wiley +1 more source
L’article présente quelques éléments de la procédure mise en place pour traiter un corpus écrit comportant 617 textes (près de 500 000 mots) relatifs aux eurorégions.
Hermand Marie-Hélène +1 more
doaj +1 more source
The discriminative lexicon is introduced as a mathematical and computational model of the mental lexicon. This novel theory is inspired by word and paradigm morphology but operationalizes the concept of proportional analogy using the mathematics of linear algebra.
R. Harald Baayen +4 more
wiley +1 more source
Characterising text mining: a systematic mapping review of the Portuguese language
Documents written in natural language constitute a major part of the artefacts produced during the software engineering life cycle. Studies indicate that more than 80% of enterprise data is stored in some sort of unstructured form, mainly as text. Therefore, the growth of user‐generated content, especially from social media, provides a huge amount of ...
Ellen Souza +8 more
wiley +1 more source
Identifying e‐Commerce in Enterprises by means of Text Mining and Classification Algorithms
Monitoring specific features of the enterprises, for example, the adoption of e‐commerce, is an important and basic task for several economic activities. This type of information is usually obtained by means of surveys, which are costly due to the amount of personnel involved in the task.
Gianpiero Bianchi +3 more
wiley +1 more source
SCAP-TT: Tagging and lemmatising Spanish tourism discourse, and beyond [PDF]
In this research note we report on the first results of SCAP, the Spanish Corpus Annotation Project, applied to tourism discourse (SCAP_tur). In particular, we present and assess a new TreeTagger parameter set for Spanish (SCAP-TT), which has been ...
Patrick Goethals +2 more
doaj

