Results 101 to 110 of about 136,207 (260)
Variations autour de tf idf et du moteur Lucene [PDF]
A l'aide d’un corpus écrit en langue française et composé de 299 requêtes, cet article analyse et compare l’efficacité du dépistage de diverses stratégies d’indexation et de recherche basées sur le modèle classique « tf idf ».
Savoy, Jacques, Dolamic, Ljiljana
core
A Country That Never Sleeps? A Web Scrapping Analysis of the 24‐h Economy Policy in Ghana
ABSTRACT In light of revitalizing Ghana's economic landscape through sustainable job creation underpinned by 24‐h operations across all key sectors, the National Democratic Congress proposed the ‘24‐h economy’ policy proposal. This study employs the web‐scraping technique through text mining and python codes to analyse 1820 comments from Facebook, X ...
Pius Gamette +3 more
wiley +1 more source
Improving a tf-idf weighted document vector embedding
We examine a number of methods to compute a dense vector embedding for a document in a corpus, given a set of word vectors such as those from word2vec or GloVe. We describe two methods that can improve upon a simple weighted sum, that are optimal in the sense that they maximizes a particular weighted cosine similarity measure.
openaire +2 more sources
IMPLEMENTASI METODE TF-IDF UNTUK APLIKASI PERINGKAS DOCUMENT BASE
Sejak terjadinya revolusi digital, informasi berbasis elektronik lebih dibutuhkan dari pada informasi berbentuk padat/kertas. Padahal, informasi yang ada dalam internet tidak sepenuhnya benar.
Kristiawan, David - 145410108
core
Pembuatan Sistem Pencarian Pekerjaan Menggunakan TF-IDF [PDF]
Seiring berjalannya waktu, kehidupan manusia terus mengalami perubahan.Begitu pula untuk konteks pekerjaan manusia sebagai kegiatan untuk pemenuhan kebutuhan.
Tirtana, Arif +2 more
core
Enhancing Fake News Detection with Word Embedding: A Machine Learning and Deep Learning Approach
The widespread dissemination of fake news on social media has necessitated the development of more sophisticated detection methods to maintain information integrity.
Mutaz A. B. Al-Tarawneh +4 more
doaj +1 more source
Revenue Recognition Comparability and Analysts’ Disclosure Processing Costs
ABSTRACT I examine whether the FASB's revenue recognition guidance under ASC 606 influences revenue comparability across firms and industries and whether revenue comparability reduces analysts’ disclosure processing costs. I extract firms’ revenue policy disclosures from 10‐K filings to measure their textual similarity and compare revenue policies ...
ANDREA TILLET
wiley +1 more source
IMPLEMENTASI EKSTRAKSI FITUR PADA PENGOLAHAN DOKUMEN BERBAHASA INDONESIA
Ekstraksi fitur merupakan proses untuk mencari nilai-nilai fitur yang terkandung dalam dokumen untuk proses text mining. Ekstraksi fitur menjadi bagian yang sangat penting dalam pengolahan dokumen pada mesin pencari karena sangat menentukan keberhasilan ...
Putu Manik Prihatini
doaj
Machine Learning-Based Hate Speech Detection in the Kazakh Language
Modern text data processing and classification methods require extensive use of machine learning and neural networks. Categorizing text into different classes has become a crucial task in many fields.
Milana Bolatbek +2 more
doaj +1 more source
Demand Estimation with Text and Image Data
ABSTRACT We propose a demand estimation approach that leverages unstructured data to infer substitution patterns. Using pre‐trained deep learning models, we extract embeddings from product images and textual descriptions and incorporate them into a mixed logit demand model.
Giovanni Compiani +2 more
wiley +1 more source

