Results 1 to 10 of about 5,103 (158)

A corpus of Chinese word segmentation agreement [PDF]

open access: yesBehavior Research Methods
Abstract The absence of explicit word boundaries is a distinctive characteristic of Chinese script, setting it apart from most alphabetic scripts, leading to word boundary disagreement among readers. Previous studies have examined how this feature may influence reading performance. However, further investigations are required to generate more
Yiu Kei Tsang, Jinger Pan
exaly   +4 more sources

A hybrid Chinese word segmentation model for quality management-related texts based on transfer learning. [PDF]

open access: yesPLoS ONE, 2022
Text information mining is a key step to data-driven automatic/semi-automatic quality management (QM). For Chinese texts, a word segmentation algorithm is necessary for pre-processing since there are no explicit marks to define word boundaries.
Peihan Wen, Linhan Feng, Tian Zhang
doaj   +2 more sources

CWSXLNet: A Sentiment Analysis Model Based on Chinese Word Segmentation Information Enhancement

open access: yesApplied Sciences, 2023
This paper proposed a method for improving the XLNet model to address the shortcomings of segmentation algorithm for processing Chinese language, such as long sub-word lengths, long word lists and incomplete word list coverage.
Shiqian Guo   +4 more
doaj   +3 more sources

The Influence of Contextual Predictability on Word Segmentation in Chinese Reading: An Eye-Tracking Study [PDF]

open access: yesBehavioral Sciences
Word segmentation is a fundamental component of lexical processing, and Chinese reading—lacking inter-word spacing—requires readers to identify word boundaries based on prior experience.
Mengchuan Song   +3 more
doaj   +2 more sources

BERTCWS: unsupervised multi-granular Chinese word segmentation based on a BERT method for the geoscience domain

open access: yesAnnals of GIS, 2023
Unlike alphabet-based languages such as English, the Chinese language has no specifying word boundaries. Segmentation, particularly for the Chinese language, is a fundamental step towards Chinese text processing, information retrieval, and knowledge ...
Qinjun Qiu, Zhong Xie, Kai Ma, Miao Tian
doaj   +3 more sources

Capsules Based Chinese Word Segmentation for Ancient Chinese Medical Books

open access: yesIEEE Access, 2018
Neural network models are popularly used in Chinese word segmentation task. The capsule architecture is proposed recently which has solved some defects of convolutional neural network. In this paper, we first introduce the capsule architecture to Chinese
Si Li   +5 more
doaj   +3 more sources

The Impact of Color Cues on Word Segmentation by L2 Chinese Readers: Evidence from Eye Movements [PDF]

open access: yesBehavioral Sciences
Chinese lacks explicit word boundary markers, creating frequent temporary segmental ambiguities where character sequences permit multiple plausible lexical analyses. Skilled native (L1) Chinese readers resolve these ambiguities efficiently.
Lin Li   +3 more
doaj   +2 more sources

Hybrid Feature Fusion Learning Towards Chinese Chemical Literature Word Segmentation

open access: yesIEEE Access, 2021
The rapid increase in the number of chemical science literature has brought challenges to researchers in search and data analysis. For many chemical scientific literature, extracting information from text and using knowledge is the focus of research ...
Xiang Li   +4 more
doaj   +1 more source

Do Chinese readers follow the national standard rules for word segmentation during reading? [PDF]

open access: yesPLoS ONE, 2013
We conducted a preliminary study to examine whether Chinese readers' spontaneous word segmentation processing is consistent with the national standard rules of word segmentation based on the Contemporary Chinese language word segmentation specification ...
Ping-Ping Liu   +3 more
doaj   +1 more source

Adaptive Chinese word segmentation [PDF]

open access: yesProceedings of the 42nd Annual Meeting on Association for Computational Linguistics - ACL '04, 2004
This paper presents a Chinese word segmentation system which can adapt to different domains and standards. We first present a statistical framework where domain-specific words are identified in a unified approach to word segmentation based on linear models. We explore several features and describe how to create training data by sampling.
Jianfeng Gao   +6 more
openaire   +2 more sources

Home - About - Disclaimer - Privacy