Results 171 to 180 of about 4,133 (219)

A double index for assessing symptoms in psychosis based on acoustic and semantic information

open access: yes
Kirdun M   +18 more
europepmc   +1 more source

Improving FastText with inverse document frequency of subwords

Pattern Recognition Letters, 2020
Abstract Word embedding is important in natural language processing, and word2vec is known as a representative algorithm. However, word2vec and many other dictionary-based word embedding algorithms create word vectors only for words that appear in the training data, ignoring morphological features of these words. The FastText algorithm was previously
Sang-Woong Lee
exaly   +2 more sources

Malware Detection and Classification Using fastText and BERT

2021 9th International Symposium on Digital Forensics and Security (ISDFS), 2021
Among the types of cyber-attacks, malware that causes high financial losses for institutions and individuals is the biggest threat to computer systems. Kinds of malware increase day-by-day and new types are released, which can easily infect our computers through injection vectors such as e-mail, websites, web applications that we use constantly.
Salih Yesir, Ibrahim Sogukpinar
openaire   +2 more sources

A Deep Investigation into fastText

2019 IEEE 21st International Conference on High Performance Computing and Communications; IEEE 17th International Conference on Smart City; IEEE 5th International Conference on Data Science and Systems (HPCC/SmartCity/DSS), 2019
The researches on text classification range from traditional machine learning methods which depend on handcrafted features to recent deep learning models which learn complex and non-linear relationships automatically. While neural models have been proven to be effective for text classification, they operate like a black box and offer no ...
Jincheng Xu, Qingfeng Du
openaire   +2 more sources

Detection and classification of malware based on FastText

2020 IEEE International Conference on Artificial Intelligence and Information Systems (ICAIIS), 2020
Nowadays, the Internet has penetrated into every corner of people's lives. It brings convenience to my life as well as certain risks. Millions of new types of malware appear every day, affecting thousands or even millions of home computer users. And attackers can use fully automated design and reuse malware, which makes the threshold for cybercrime ...
Luming Feng, Yanpeng Cui, Jianwei Hu
openaire   +1 more source

Text Classification Model Based on fastText

2020 IEEE International Conference on Artificial Intelligence and Information Systems (ICAIIS), 2020
Most text classification models based on traditional machine learning algorithms have problems such as curse of dimensionality and poor performance. In order to solve the above problems, this paper proposes a text classification model based on fastText.
Tengjun Yao, Zhengang Zhai, Bingtao Gao
openaire   +1 more source

Subwords-Only Alternatives to fastText for Morphologically Rich Languages

Programming and Computer Software, 2021
In this work, we present purely subword-based alternatives to fastText word embedding algorithm The alternatives are modifications of the original fastText model, but rely on subword information only, eliminating the reliance on word-level vectors and at the same time helping to dramatically reduce the size of embeddings.
Tsolak Ghukasyan   +2 more
openaire   +2 more sources

Detecting Webshell Based on Random Forest with FastText

Proceedings of the 2018 International Conference on Computing and Artificial Intelligence, 2018
Web-based remote access Trojan (or webshell) is a kind of tool for network intrusion, which can be uploaded to a website to access web service management authority. Once attacker injected successfully, it can cause great damage so that it is crucial to detect webshell effectively. Webshells are flexible and changeable by using of obfuscation techniques,
Yong Fang 0002   +3 more
openaire   +2 more sources

Home - About - Disclaimer - Privacy