Results 111 to 120 of about 645,699 (249)
Domain Adaptation in Sequence Labelling: A Case Study for Two South African Languages
In this paper, we investigate domain adaptation for Part-of-speech (POS) tagging of two under-resourced South African languages, isiZulu and Sesotho sa Leboa, by studying its effect on the POS tagging results and how to possibly predict what quality can
Tanja Gaustad, Roald Eiselen
doaj +1 more source
Recent Advances in Text Anonymization: A Systematic Review
This survey presents a unified, cross‐domain overview of text anonymization from 2021 to 2025, covering methods from rule‐based to LLM‐based approaches, relevant datasets, benchmarks, and evaluation strategies. It highlights the critical trade‐off between privacy and utility, underscores the importance of standardizing evaluation protocols and metrics,
Marina Litvak, Alípio Jorge
wiley +1 more source
Hybrid POS tagging with generalized unknown-word handling [PDF]
This paper presents POSTAG 1 as a statistical/rulebased hybrid part-of-speech (POS) tagging system with generalized unknown-word handling. The POSTAG integrates morphological analysis with statistical POS disambiguation and post rule-based error ...
Jong-hyeok Lee +2 more
core
Ten Pairs to Tag – Multilingual POS Tagging via Coarse Mapping between Embeddings
In the absence of annotations in the target language, multilingual models typically draw on extensive parallel resources. In this paper, we demonstrate that accurate multilingual part-of-speech (POS) tagging can be done with just a few (e.g., ten) word ...
Yuan Zhang +3 more
semanticscholar +1 more source
ABSTRACT Online dissemination of conspiracy theories (CTs) during epidemics poses significant risks to public health. This paper addresses the problem of detecting CTs in social media posts with an emphasis on the resource‐constrained scenarios characterized by the scarcity of labelled datasets and expert annotations, and the lack of computational ...
Ipek Baris Schlicht +4 more
wiley +1 more source
A comparative study on Arabic POS tagging using Quran corpus [PDF]
POS tagging is the process of computationally assigning correct part of speech to each word of a given input text depending on the context. Different POS tagging techniques in the literature have been developed and experimented mostly for English ...
Alashqar, Abdelkareem M
core
USzeged: Identifying Verbal Multiword Expressions with POS Tagging and Parsing Techniques
The paper describes our system submitted for the Workshop on Multiword Expressions’ shared task on automatic identification of verbal multiword expressions.
K. Simkó, V. Kovács, V. Vincze
semanticscholar +1 more source
The double modal construction in English world wide
Abstract The dual foci of the present study of double modals are their semantic characteristics and their distribution across regional varieties of English world wide. Tokens were extracted from GloWbE:Blogs, a database whose great size and informal tenor facilitated the investigation of this low‐frequency non‐standard feature. Double modals were found
Peter Collins, Adam Smith
wiley +1 more source
An In-depth Study on POS Tagging for Assamese Language [PDF]
: An automatic POS tagger is very essential component of any Natural Language Processing (NLP) work. It is one of the important steps towards the processing of Natural Language.
Boro, Dr. Karabi Kherkatary +1 more
core
A Data-Driven Model for Automated Chinese Word Segmentation and POS Tagging. [PDF]
Xu Q, Wang Z.
europepmc +1 more source

