Grammatical error correction for low-resource languages: a review of challenges, strategies, computational and future directions. [PDF]
Marier SM, Chen X, Zhu L, Kong X.
europepmc +1 more source
Spanish-language text classification for environmental evidence synthesis using multilingual pre-trained models. [PDF]
Berdejo-Espinola V +4 more
europepmc +1 more source
Towards a resource for multilingual lexicons: an MT assisted and human-in-the-loop multilingual parallel corpus with multi-word expression annotation. [PDF]
Han L +5 more
europepmc +1 more source
IGAR: Indonesian government applications review for sentiment analysis dataset. [PDF]
Isnan M, Pardamean B.
europepmc +1 more source
AsanteTwiSenti: A Sentiment dataset of Ghanaian Asante Twi Tweets in a multilingual context. [PDF]
Gyimah MS +5 more
europepmc +1 more source
Adding LLMs to the psycholinguistic norming toolbox: A practical guide to getting the most out of human ratings. [PDF]
Conde J +9 more
europepmc +1 more source
PolCSBD: A contextual dataset of code-mixed Bengali and Banglish for political counter-speech detection. [PDF]
Mehedi CM +5 more
europepmc +1 more source
Key concept learning for medical vision language model with reasoning capabilities. [PDF]
Lou W +7 more
europepmc +1 more source
Birth of a Language in the Backlands of Brazil. [PDF]
Almeida-Silva A +4 more
europepmc +1 more source
A dataset of genuine and synthetic Portuguese speech covering European, Brazilian, African, and regional accents - <i>An open resource for speech deepfake detection and trustworthy AI research</i>. [PDF]
Oliveira AMV +3 more
europepmc +1 more source

