Results 261 to 270 of about 14,339,869 (296)
Some of the next articles are maybe not open access.
ACM SIGIR Forum, 2011
The Web N-gram Workshop was held on July 23, 2010 in Geneva, Switzerland, in conjunction with the 33rd Annual ACM SIGIR Conference. The workshop brought together leaders in information retrieval and language modeling to discuss the challenges in information retrieval and how language modeling approaches may help address some of these challenges, with a
Chengxiang Zhai +4 more
openaire +1 more source
The Web N-gram Workshop was held on July 23, 2010 in Geneva, Switzerland, in conjunction with the 33rd Annual ACM SIGIR Conference. The workshop brought together leaders in information retrieval and language modeling to discuss the challenges in information retrieval and how language modeling approaches may help address some of these challenges, with a
Chengxiang Zhai +4 more
openaire +1 more source
A succinct N-gram language model
Proceedings of the ACL-IJCNLP 2009 Conference Short Papers on - ACL-IJCNLP '09, 2009Efficient processing of tera-scale text data is an important research topic. This paper proposes lossless compression of N-gram language models based on LOUDS, a succinct data structure. LOUDS succinctly represents a trie with M nodes as a 2M + 1 bit string. We compress it further for the N-gram language model structure.
Taro Watanabe +2 more
openaire +2 more sources
Distributing N-Gram Graphs for Classification
2017N-gram models have been an established choice for language modeling in machine translation, summarization and other tasks. Recently n-gram graphs managed to capture significant language characteristics that go beyond mere vocabulary and grammar, for tasks such as text classification.
Ioannis Kontopoulos +2 more
openaire +1 more source
An Indic Language N-gram Viewer
Proceedings of the 8th Annual Meeting of the Forum for Information Retrieval Evaluation, 2016In this paper, we introduce culturomic studies on diachronic corpora for three languages -- Hindi, Bengali, and English, in the same spirit as that of the Google Culturomics Project [13]. Based on Hindi, Bengali, and English text from news blogs and newspapers, we are able to extract trajectories of salient words that are of importance in contemporary ...
Shanta Phani +3 more
openaire +1 more source
2019
In this and the following chapters, we present two ideas related to the non-linear construction of n-grams. Recall that the non-linear construction consists in taking the elements which form n-grams in a different order than the surface (textual) representation, i.e., in a different way than words (lemmas, POS tags, etc.) appear in a text.
openaire +1 more source
In this and the following chapters, we present two ideas related to the non-linear construction of n-grams. Recall that the non-linear construction consists in taking the elements which form n-grams in a different order than the surface (textual) representation, i.e., in a different way than words (lemmas, POS tags, etc.) appear in a text.
openaire +1 more source
2014
A statistical language model defines a probability distribution over a set of symbol sequences from some finite inventory. An especially simple yet very powerful concept for the formal description of statistical language models is formed by their representation using Markov chains or so-called n-gram models.
openaire +1 more source
A statistical language model defines a probability distribution over a set of symbol sequences from some finite inventory. An especially simple yet very powerful concept for the formal description of statistical language models is formed by their representation using Markov chains or so-called n-gram models.
openaire +1 more source
Proceedings of the 7th ACM international conference on Web search and data mining, 2014
We propose a Bayesian nonparametric topic model that rep- resents relationships between given labels and the corre- sponding words/phrases, from supervised articles. Unlike existing supervised topic models, our proposal, supervised N-gram topic model (SNT), focuses on both a number of topics and power-law distribution in the word frequencies to extract
openaire +1 more source
We propose a Bayesian nonparametric topic model that rep- resents relationships between given labels and the corre- sponding words/phrases, from supervised articles. Unlike existing supervised topic models, our proposal, supervised N-gram topic model (SNT), focuses on both a number of topics and power-law distribution in the word frequencies to extract
openaire +1 more source
2019
Another idea related to the non-linear construction of n-grams, i.e., using distinct elements or distinct order of their appearance in a text, is the idea of replacing words by their synonyms or by the generalized concepts that correspond to the words according to a certain ontology.
openaire +1 more source
Another idea related to the non-linear construction of n-grams, i.e., using distinct elements or distinct order of their appearance in a text, is the idea of replacing words by their synonyms or by the generalized concepts that correspond to the words according to a certain ontology.
openaire +1 more source
Biofilm patterns in gram-positive and gram-negative bacteria
Microbiological Research, 2021Rohit Ruhal
exaly
Part-of-speech n-gram and word n-gram fused language model
6th European Conference on Speech Communication and Technology, 1999Hirofumi Yamamoto, Yoshinori Sagisaka
openaire +2 more sources

