Results 21 to 30 of about 4,133 (219)

BERT and fastText Embeddings for Automatic Detection of Toxic Speech [PDF]

open access: yes2020 International Multi-Conference on: “Organization of Knowledge and Advanced Technologies” (OCTA), 2020
With the expansion of Internet usage, catering to the dissemination of thoughts and expressions of an individual, there has been an immense increase in the spread of online hate speech. Social media, community forums, discussion platforms are few examples of common playground of online discussions where people are freely allowed to communicate. However,
Ashwin Geet D'Sa   +2 more
openaire   +2 more sources

FastText Embedding for Mossi

open access: yes, 2022
FastText Embedding for Mossi, as part of the PhD thesis of David Ifeoluwa ...
David Ifeoluwa Adelani
core   +1 more source

Modern Approaches to Detect and Classify Comment Toxicity Using Neural Networks

open access: yesМоделирование и анализ информационных систем, 2020
The growth of popularity of online platforms which allow users to communicate with each other, share opinions about various events, and leave comments boosted the development of natural language processing algorithms. Tens of millions of messages per day
Sergey V. Morzhov
doaj   +1 more source

FastText Embedding for Ghomala

open access: yes, 2022
FastText Embedding for Ghomala, as part of the PhD thesis of David Ifeoluwa ...
David Ifeoluwa Adelani
core   +1 more source

Comparison of the accuracy of Japanese synonym identifications using word embeddings in the radiological technology field

open access: yesScientific Reports, 2023
The terminology in radiological technology is crucial, encompassing a broad range of principles from radiation to medical imaging, and involving various specialists. This study aimed to evaluate the accuracy of automatic synonym detection considering the
Ayako Yagahara, Noriya Yokohama
doaj   +1 more source

FastText Embedding for Shona

open access: yes, 2022
FastText Embedding for Shona, as part of the PhD thesis of David Ifeoluwa ...
David Ifeoluwa Adelani
core   +1 more source

Word-embedding Based Text Vectorization Using Clustering

open access: yesМоделирование и анализ информационных систем, 2021
It is known that in the tasks of natural language processing, the representation of texts by vectors of fixed length using word-embedding models makes sense in cases where the vectorized texts are short.The longer the texts being compared, the worse the ...
Vitaly I. Yuferev, Nikolai A. Razin
doaj   +1 more source

FastText embeddings for Galician

open access: yes, 2021
FastText embeddings for Galician (skip-gram, 300 dimensions, window size of 5 ...
Anonymous
core   +1 more source

Neural Language Models for Nineteenth-Century English

open access: yesJournal of Open Humanities Data, 2021
We present four types of neural language models trained on a large historical dataset of books in English, published between 1760 and 1900, and comprised of ≈5.1 billion tokens.
Kasra Hosseini   +3 more
doaj   +1 more source

FastText Embedding for Hausa

open access: yes, 2022
FastText Embedding for Hausa, as part of the PhD thesis of David Ifeoluwa ...
David Ifeoluwa Adelani
core   +1 more source

Home - About - Disclaimer - Privacy