Results 1 to 10 of about 2,332 (184)

Large language models unlock large text corpora in the search for data on medicinal plants and fungi

open access: yesPLANTS, PEOPLE, PLANET, EarlyView.
Millions of people worldwide depend on plants and fungi for their healthcare needs, yet our understanding of which species are used medicinally—and how—remains fragmented and incomplete. As part of Kew's Plants for Health service, we developed a pipeline to mine large text corpora for information on natural products, initially focusing on the medicinal
Adam Richard‐Bollans   +6 more
wiley   +1 more source

WordNet [PDF]

open access: yesCommunications of the ACM, 1992
Because meaningful sentences are composed of meaningful words, any system that hopes to process natural languages as people do must have information about words and their meanings. This information is traditionally provided through dictionaries, and machine-readable dictionaries are now widely available.
openaire   +6 more sources

Measuring New Venture Legitimacy: A Multi‐dimensional Approach Using Computational Textual Analysis

open access: yesBritish Journal of Management, EarlyView.
Abstract Legitimacy is critical for the growth, survival and performance of ventures, yet efforts to measure legitimacy empirically typically focus on a single dimension of this multi‐dimensional construct. In this article, we develop a multi‐dimensional framework for measuring legitimacy using computational text analysis.
Ali Ghods   +3 more
wiley   +1 more source

Utemeljevanje sloWNeta na korpusnih podatkih

open access: yesSlovenščina 2.0: Empirične, aplikativne in interdisciplinarne raziskave, 2013
Wordnet lahko izdelamo na podlagi že obstoječega tujejezičnega wordneta ali pa kot osnovo za gradnjo vzamemo korpusne podatke. Prvi pristop je preprostejši in enostavnejši, zaradi česar ga razvijalci tudi najpogosteje uporabljajo.
Darja Fišer   +2 more
doaj   +1 more source

Persian Wordnet Construction using Supervised Learning

open access: yesInternational Journal of Information and Communication Technology Research, 2017
This paper presents an automated supervised method for Persian wordnet construction. Using a Persian corpus and a bi-lingual dictionary, the initial links between Persian words and Princeton WordNet synsets have been generated.
Zahra Mousavi   +2 more
doaj  

Lexicogrammatical Features Predicting Grade‐Based Outcomes in a Post‐Secondary Content Course: An NLP Approach

open access: yesInternational Journal of Applied Linguistics, EarlyView.
ABSTRACT This exploratory study investigated lexicogrammatical features associated with grade‐based differences in a corpus of 304 business papers taken from an upper‐level undergraduate course at a large North American university. Using five natural language processing tools, 719 candidate linguistic variables were analyzed for their relationship with
Randy Appel   +2 more
wiley   +1 more source

Sentiment Analysis of Imbalanced Dataset Through Data Augmentation and Generative Annotation Using DistilBERT and Low‐Rank Fine‐Tuning

open access: yesApplied AI Letters, Volume 7, Issue 3, October 2026.
The proposed model takes an imbalanced English dataset as input and balances it by Generating Synthetic Tweets (GST) using GPT‐based paraphrasing. These new tweets are then annotated with one of 10 categorical labels representing their primary meaning, using another large language model like GPT.
Hossein Nekkouei Nasrabadi   +1 more
wiley   +1 more source

3DText: Perceiving sentence-level text on 3-D model of emotions

open access: yesEngineering and Applied Science Research, 2020
Emotion is a psychological process which reveals the sentiments and feelings of a human being. Relating emotions detection process with psychological theory of emotions serves as the strong foundation for the system.
Shalu Choudhary, Shikha Jain
doaj   +1 more source

Tracking Climate and Environmental Attention: A News‐Based Composite Index

open access: yesCorporate Social Responsibility and Environmental Management, Volume 33, Issue 5, Page 7095-7123, September 2026.
ABSTRACT This study introduces the Climate and Environmental Attention Index, a composite indicator that tracks media attention to climate and environmental issues. Based on the Semantic Brand Score, the proposed index extracts significant signals from unstructured text, going beyond traditional measures of word frequency and sentiment.
Gianna Figà‐Talamanca   +3 more
wiley   +1 more source

The data-driven Bulgarian WordNet: BTBWN

open access: yesCognitive Studies | Études cognitives, 2018
The data-driven Bulgarian WordNet: BTBWN The paper presents our work towards the simultaneous creation of a data-driven WordNet for Bulgarian and a manually annotated treebank with semantic information.
Petya Osenova, Kiril Simov
doaj   +1 more source

Home - About - Disclaimer - Privacy