Results 11 to 20 of about 160 (101)

TEICORPO: A Conversion Tool for Spoken Language Transcription with a Pivot File in TEI

open access: yesJournal of the Text Encoding Initiative, 2021
CORLI is a consortium of Huma-Num, the French national infrastructure dedicated to the technical support and promotion of digital humanities. The goal of CORLI is to promote and provide tools and information for good and efficient research practices in ...
Christophe Parisse   +2 more
doaj   +1 more source

Neural POS tagging of shahmukhi by using contextualized word representations

open access: yesJournal of King Saud University: Computer and Information Sciences, 2023
Part of Speech (POS) tagging has a preliminary role in building natural language processing applications. This paper presents the development and evaluation of the first POS tagged corpus along with a Bi-directional long-short memory (BiLSTM) network ...
Amina Tehseen   +4 more
doaj   +1 more source

Formalización de reglas para la detección del plural en castellano en el caso de unidades no diccionarizadas

open access: yesLinguamática, 2020
En este artículo ofrecemos una formalización de reglas de pluralización en castellano para ser utilizada concretamente en el procesamiento de términos especializados, ya que con frecuencia estos no se encuentran registrados en los diccionarios de lengua ...
Rogelio Nazar, Amparo Galdames
doaj   +1 more source

Noun phrase complexity in Ghanaian English

open access: yesWorld Englishes, Volume 39, Issue 3, Page 377-393, September 2020., 2020
Abstract This study compares the complexity of the noun phrase (NP) in Ghanaian English in a real‐time perspective. Based on the Historical Corpus English in Ghana (1966–1975) and the Ghanaian component of the International Corpus of English (mainly 2000s), representing the early and late stages of structural nativisation in the dynamic model, NP ...
Thorsten Brato
wiley   +1 more source

Traitement de données issues d’un corpus écrit multilingue. Approche agile pour l’analyse du discours eurorégional

open access: yesSHS Web of Conferences, 2015
L’article présente quelques éléments de la procédure mise en place pour traiter un corpus écrit comportant 617 textes (près de 500 000 mots) relatifs aux eurorégions.
Hermand Marie-Hélène   +1 more
doaj   +1 more source

The Discriminative Lexicon: A Unified Computational Model for the Lexicon and Lexical Processing in Comprehension and Production Grounded Not in (De)Composition but in Linear Discriminative Learning

open access: yesComplexity, Volume 2019, Issue 1, 2019., 2019
The discriminative lexicon is introduced as a mathematical and computational model of the mental lexicon. This novel theory is inspired by word and paradigm morphology but operationalizes the concept of proportional analogy using the mathematics of linear algebra.
R. Harald Baayen   +4 more
wiley   +1 more source

Characterising text mining: a systematic mapping review of the Portuguese language

open access: yesIET Software, Volume 12, Issue 2, Page 49-75, April 2018., 2018
Documents written in natural language constitute a major part of the artefacts produced during the software engineering life cycle. Studies indicate that more than 80% of enterprise data is stored in some sort of unstructured form, mainly as text. Therefore, the growth of user‐generated content, especially from social media, provides a huge amount of ...
Ellen Souza   +8 more
wiley   +1 more source

Identifying e‐Commerce in Enterprises by means of Text Mining and Classification Algorithms

open access: yesMathematical Problems in Engineering, Volume 2018, Issue 1, 2018., 2018
Monitoring specific features of the enterprises, for example, the adoption of e‐commerce, is an important and basic task for several economic activities. This type of information is usually obtained by means of surveys, which are costly due to the amount of personnel involved in the task.
Gianpiero Bianchi   +3 more
wiley   +1 more source

SCAP-TT: Tagging and lemmatising Spanish tourism discourse, and beyond [PDF]

open access: yesIbérica, 2017
In this research note we report on the first results of SCAP, the Spanish Corpus Annotation Project, applied to tourism discourse (SCAP_tur). In particular, we present and assess a new TreeTagger parameter set for Spanish (SCAP-TT), which has been ...
Patrick Goethals   +2 more
doaj  

Home - About - Disclaimer - Privacy