Results 41 to 50 of about 27,036 (297)
In natural language, the phenomenon of polysemy is widespread, which makes it very difficult for machines to process natural language. Word sense disambiguation is a key issue in the field of natural language processing.
Lei Wang, Qun Ai
doaj +1 more source
Word-based largest chunks for Agreement Groups processing: Cross-linguistic observations
The present study reports results from a series of computer experiments seeking to combine word-based Largest Chunk (LCh) segmentation and Agreement Groups (AG) sequence processing.
László Drienkó
doaj +1 more source
Adaptive Chinese word segmentation [PDF]
This paper presents a Chinese word segmentation system which can adapt to different domains and standards. We first present a statistical framework where domain-specific words are identified in a unified approach to word segmentation based on linear models. We explore several features and describe how to create training data by sampling.
Jianfeng Gao +6 more
openaire +1 more source
Early syllabic segmentation of fluent speech by infants acquiring French. [PDF]
Word form segmentation abilities emerge during the first year of life, and it has been proposed that infants initially rely on two types of cues to extract words from fluent speech: Transitional Probabilities (TPs) and rhythmic units.
Louise Goyet +2 more
doaj +1 more source
Word Segmentation as Graph Partition
We propose a new approach to the Chinese word segmentation problem that considers the sentence as an undirected graph, whose nodes are the characters. One can use various techniques to compute the edge weights that measure the connection strength between characters.
Yuanhao Liu, Sheng Yu 0002
openaire +2 more sources
Detecting “protein words” through unsupervised word segmentation [PDF]
Unsupervised word segmentation methods were applied to analyze protein sequences. Protein sequences, such as “MTMDKSELVQKA…,” were used as input to these methods. Segmented protein word sequences, such as “MTM DKSE LVQKA,” were then obtained.
Liang Wang, Kaiyong Zhao
openaire +2 more sources
A summary of the 2012 JHU CLSP Workshop on Zero Resource Speech Technologies and Models of Early Language Acquisition [PDF]
We summarize the accomplishments of a multi-disciplinary workshop exploring the computational and scientific issues surrounding zero resource (unsupervised) speech technologies and related models of early language acquisition.
Johnson, Mark +53 more
core +1 more source
Method of Word Segmentation in Laos Based on Maximal Matching of Syllables
Word segmentation is an important support of semantic analysis, Machine Translation, QA, knowledge mapping research work, mainly used in information retrieval, text processing, data processing and many other areas of Natural Language Processing ...
Huo Wenjie +3 more
doaj +1 more source
Experience with a second language affects the use of fundamental frequency in speech segmentation. [PDF]
This study investigates whether listeners' experience with a second language learned later in life affects their use of fundamental frequency (F0) as a cue to word boundaries in the segmentation of an artificial language (AL), particularly when the cues ...
Annie Tremblay +7 more
doaj +1 more source
Unlike alphabet-based languages such as English, the Chinese language has no specifying word boundaries. Segmentation, particularly for the Chinese language, is a fundamental step towards Chinese text processing, information retrieval, and knowledge ...
Qinjun Qiu, Zhong Xie, Kai Ma, Miao Tian
doaj +1 more source

