Results 1 to 10 of about 2,332 (184)
Large language models unlock large text corpora in the search for data on medicinal plants and fungi
Millions of people worldwide depend on plants and fungi for their healthcare needs, yet our understanding of which species are used medicinally—and how—remains fragmented and incomplete. As part of Kew's Plants for Health service, we developed a pipeline to mine large text corpora for information on natural products, initially focusing on the medicinal
Adam Richard‐Bollans +6 more
wiley +1 more source
Because meaningful sentences are composed of meaningful words, any system that hopes to process natural languages as people do must have information about words and their meanings. This information is traditionally provided through dictionaries, and machine-readable dictionaries are now widely available.
openaire +6 more sources
Measuring New Venture Legitimacy: A Multi‐dimensional Approach Using Computational Textual Analysis
Abstract Legitimacy is critical for the growth, survival and performance of ventures, yet efforts to measure legitimacy empirically typically focus on a single dimension of this multi‐dimensional construct. In this article, we develop a multi‐dimensional framework for measuring legitimacy using computational text analysis.
Ali Ghods +3 more
wiley +1 more source
Utemeljevanje sloWNeta na korpusnih podatkih
Wordnet lahko izdelamo na podlagi že obstoječega tujejezičnega wordneta ali pa kot osnovo za gradnjo vzamemo korpusne podatke. Prvi pristop je preprostejši in enostavnejši, zaradi česar ga razvijalci tudi najpogosteje uporabljajo.
Darja Fišer +2 more
doaj +1 more source
Persian Wordnet Construction using Supervised Learning
This paper presents an automated supervised method for Persian wordnet construction. Using a Persian corpus and a bi-lingual dictionary, the initial links between Persian words and Princeton WordNet synsets have been generated.
Zahra Mousavi +2 more
doaj
ABSTRACT This exploratory study investigated lexicogrammatical features associated with grade‐based differences in a corpus of 304 business papers taken from an upper‐level undergraduate course at a large North American university. Using five natural language processing tools, 719 candidate linguistic variables were analyzed for their relationship with
Randy Appel +2 more
wiley +1 more source
The proposed model takes an imbalanced English dataset as input and balances it by Generating Synthetic Tweets (GST) using GPT‐based paraphrasing. These new tweets are then annotated with one of 10 categorical labels representing their primary meaning, using another large language model like GPT.
Hossein Nekkouei Nasrabadi +1 more
wiley +1 more source
3DText: Perceiving sentence-level text on 3-D model of emotions
Emotion is a psychological process which reveals the sentiments and feelings of a human being. Relating emotions detection process with psychological theory of emotions serves as the strong foundation for the system.
Shalu Choudhary, Shikha Jain
doaj +1 more source
Tracking Climate and Environmental Attention: A News‐Based Composite Index
ABSTRACT This study introduces the Climate and Environmental Attention Index, a composite indicator that tracks media attention to climate and environmental issues. Based on the Semantic Brand Score, the proposed index extracts significant signals from unstructured text, going beyond traditional measures of word frequency and sentiment.
Gianna Figà‐Talamanca +3 more
wiley +1 more source
The data-driven Bulgarian WordNet: BTBWN
The data-driven Bulgarian WordNet: BTBWN The paper presents our work towards the simultaneous creation of a data-driven WordNet for Bulgarian and a manually annotated treebank with semantic information.
Petya Osenova, Kiril Simov
doaj +1 more source

