Constructing the Corpus of Children's Video Media (CCVM): A new resource and guidelines for constructing comparable and reusable corpora. [PDF]
Gowenlock A +3 more
europepmc +1 more source
Writing creativity, cohesion, and formal linguistic competence in LLMs: A comparative evaluation based on English and Chinese continuation writing. [PDF]
Zhang Y +5 more
europepmc +1 more source
On the Design of a Sign Language Corpus of Medical Terms for Automatic Translation Systems: Mixed Methods Approach. [PDF]
Marcolino MS +3 more
europepmc +1 more source
Representing national images of self and others through China's diplomatic discourse: A corpus-based study. [PDF]
Pan F, Li W, Li T.
europepmc +1 more source
ANCHOLIK-NER: A benchmark dataset for Bangla regional named entity recognition. [PDF]
Paul B +7 more
europepmc +1 more source
Dataset on multiregional variations of Bangla language (BD-Dialect). [PDF]
Rahman A, Muna NH, Prity MS.
europepmc +1 more source
Enhanced extractive text summarization framework for low-resourced Urdu language. [PDF]
Nazir S +4 more
europepmc +1 more source
Learning semantic similarity from sentence pairs using hybrid features centric approach and explainable siamese neural networks. [PDF]
Zhao W, Hu C.
europepmc +1 more source
Warm realism in Chinese microdramas: A term frequency analysis of audience comments on The Puppy Laifu. [PDF]
Tian Y, Cheng P.
europepmc +1 more source

