Results 61 to 70 of about 4,133 (219)
LLM‐Assisted Topic Modelling for Hate Speech Characterization
ABSTRACT In the digital era, the internet and social media have transformed communication but have also facilitated the spread of hate speech and disinformation, leading to radicalization, polarisation and toxicity. This is especially concerning for media outlets due to their significant role in shaping public discourse. This study examines the topics,
Alejandro Buitrago López +2 more
wiley +1 more source
Morphological Skip-Gram: Replacing FastText characters n-gram with morphological knowledge
Natural language processing systems have attracted much interest of the industry. This branch of study is composed of some applications such as machine translation, sentiment analysis, named entity recognition, question and answer, and others.
Thiago Dias Bispo +3 more
doaj +1 more source
ABSTRACT The rapid expansion of multilingual digital platforms has made the accurate analysis of user‐generated content across different languages and cultural contexts increasingly essential. However, existing methods struggle to maintain consistent performance due to linguistic diversity, morphological complexity, and structural variations in text ...
Abdulkadir Şeker
wiley +1 more source
Evaluating Fasttext and Glove Embeddings for Sentiment Analysis of AI-Generated Ghibli-Style Images
The development of text-to-image generation technology based on artificial intelligence has triggered mixed public reactions, especially when applied to iconic visual styles such as Studio Ghibli.
I Gusti Ngurah Sentana Putra +5 more
doaj +1 more source
COVID-19 Tweets Classification Based on a Hybrid Word Embedding Method
In March 2020, the World Health Organisation declared that COVID-19 was a new pandemic. This deadly virus spread and affected many countries in the world.
Yosra Didi, Ahlam Walha, Ali Wali
doaj +1 more source
This study presents a comprehensive analysis of Bengali AI resources, including NLP tools, machine translation systems, speech technologies, and benchmarking datasets. The findings highlight fragmented development across tasks and emphasize the need for integrated resources, unified benchmarks, and multimodal datasets to build a scalable and robust ...
Fairooz Maliha +5 more
wiley +1 more source
Evaluasi Pengukuran Semantik Sinonim KBBI Menggunakan Pendekatan Word Embedding
Kamus Besar Bahasa Indonesia (KBBI) ialah salah satu sumber utama penyedia data dalam penelitian penentuan kemiripan makna kata dalam bahasa Indonesia.
Muhammad Rafli Aditya H. +3 more
doaj +1 more source
Comparative AI‐Based Framework for Name‐Based Demographic Inference in the Indian Subcontinent
A multi‐task framework of five AI models (SVM, XGBoost, LightGBM, BiLSTM, and XLM‐RoBERTa) infers nationality, religion, and gender from personal names across seven Indian subcontinent countries. Character‐level TF‐IDF with SVM achieves the highest accuracy: 83.23% for nationality, 92.94% for religion, and 92.67% for gender across 7581 names.
Sherin Sultana +5 more
wiley +1 more source
New Heuristics for Stable LDA Parameter Search
ABSTRACT The Latent Dirichlet Allocation (LDA) algorithm automatically extracts latent topics from a textual corpus, but configuring its parameters can be difficult and time‐consuming. Optimization algorithms can help determine the best parameters, but not necessarily the optimal ones.
Simon‐Olivier Harel +2 more
wiley +1 more source
FastText Embedding for Yoruba, as part of the PhD thesis of David Ifeoluwa Adelani. The embedding was trained for the Massive vs.
David Ifeoluwa Adelani
core +1 more source

