Results 241 to 250 of about 131,669 (281)
Some of the next articles are maybe not open access.

AIFdb Corpora

2014
This paper introduces AIFdb Corpora, an addition to AIFdb for collecting and presenting sets of AIF argument maps. AIFdb Corpora allows any user to create a corpus and share the contents.
John Lawrence, Chris Reed 0001
openaire   +3 more sources

Child-Language Corpora

2020
Together with experiments, the main method in the study of child language development is the analysis of behavioral changes of individual children over time. For this purpose recordings of infants’ and children’s naturalistic interactions in a variety of languages spoken in different cultural contexts are key.
Stoll, Sabine, Schikowski, Robert
openaire   +1 more source

Parsed corpora for linguistics

Proceedings of the EACL 2009 Workshop on the Interaction between Linguistics and Computational Linguistics Virtuous, Vicious or Vacuous? - ILCL '09, 2009
Knowledge-based parsers are now accurate, fast and robust enough to be used to obtain syntactic annotations for very large corpora fully automatically. We argue that such parsed corpora are an interesting new resource for linguists. The argument is illustrated by means of a number of recent results which were established with the help of parsed corpora.
van Noord, G.J.M., Bouma, G.
openaire   +2 more sources

On the Assessment of Text Corpora

2010
Classifier-independent measures are important to assess the quality of corpora. In this paper we present supervised and unsupervised measures in order to analyse several data collections for studying the following features: domain broadness, shortness, class imbalance, and stylometry.
David Pinto 0001   +2 more
openaire   +1 more source

Corpora

2016
In this chapter the use of corpora in natural-language processing (NLP) is overviewed. The chapter begins by defining what a corpus is. In doing so it introduces different types of corpora such as monolingual, parallel and comparable corpora. It also discusses key issues in corpus design, notably balance and representativeness.
openaire   +1 more source

Characterizing Weblog Corpora

2010
In order to exploit the huge volume of information being published in the blogosphere, it is essential to provide techniques such as clustering, which can automatically analyze and classify their contents. However these typically can produce better results when dealing with wide domain full-text documents. In most cases however, blogs can be considered
Fernando Pérez-Téllez   +3 more
openaire   +2 more sources

Interoperability of Corpora and Annotations

2012
This paper describes the application of OWL and RDF to address the interoperability of linguistic corpora and linguistic annotations within such corpora. Interoperability of linguistic corpora involves two aspects: Structural interoperability (annotations of different origin are represented using the same formalism) and conceptual interoperability ...
openaire   +1 more source

Text corpora, professional translators and translator training

Interpreter and Translator Trainer, 2022
Mikhail Mikhailov
exaly  

Freely Available Arabic Corpora: A Scoping Review

Computer Methods and Programs in Biomedicine Update, 2022
Alaa A Abd-Alrazaq   +2 more
exaly  

Home - About - Disclaimer - Privacy