Results 231 to 240 of about 10,051 (260)
Some of the next articles are maybe not open access.
2015
One of the major bottlenecks in the development of Statistical Machine Translation systems for most language pairs is the lack of bilingual parallel training data.
openaire +1 more source
One of the major bottlenecks in the development of Statistical Machine Translation systems for most language pairs is the lack of bilingual parallel training data.
openaire +1 more source
Comparative evaluation of two arabic speech corpora
Proceedings of the 6th International Conference on Natural Language Processing and Knowledge Engineering(NLPKE-2010), 2010The aim of this paper is to conduct a constructive and comparative evaluation between two important Arabic corpora for two different Arabic dialects, namely, Saudi dialect corpus that was collected by King Abdulaziz City for Science and Technology (KACST), and a Levantine Arabic dialect corpus.
Yousef Ajami Alotaibi, Ali Hamid Meftah
openaire +1 more source
Terminology Extraction from Comparable Corpora for Latvian
2012This paper presents the work on terminology extraction from comparable corpora for Latvian. In the first section we introduce our work; the second section briefly describes the concept of the project and the implemented general terminology processing chain; the following two sections focus on terminology extraction workflow for Latvian and evaluation ...
Tatiana Gornostay +5 more
openaire +1 more source
A Collection of Comparable Corpora for Under-resourced Languages
2010This paper presents work on collecting comparable corpora for 9 language pairs: Estonian-English, Latvian-English, Lithuanian-English, Greek-English, Greek-Romanian, Croatian-English, Romanian-English, Romanian-German and Slovenian-English. The objective of this work was to gather texts from the same domains and genres and with a similar level of ...
Inguna Skadina +6 more
openaire +1 more source
Building and Using Comparable Corpora for Multilingual Natural Language Processing
Synthesis Lectures on Human Language Technologies, 2023Pierre Zweigenbaum
exaly
Comparing the Level of Code-Switching in Corpora
Proceedings of the Language Resources and Evaluation Conference, 2016Björn Gambäck, Amitava Das 0001
openaire +2 more sources

