Results 21 to 30 of about 52,865,053 (238)
TACO: Topics in Algorithmic COde generation dataset
We introduce TACO, an open-source, large-scale code generation dataset, with a focus on the optics of algorithms, designed to provide a more challenging training dataset and evaluation benchmark in the field of code generation models. TACO includes competition-level programming questions that are more challenging, to enhance or evaluate problem ...
Rongao Li +8 more
openaire +3 more sources
This study constructs a comprehensive index to effectively judge the optimal number of topics in the LDA topic model. Based on the requirements for selecting the number of topics, a comprehensive judgment index of perplexity, isolation, stability, and ...
Jingxian Gan, Yong Qi
doaj +1 more source
Retrieval Topic Recurrent Memory Network for Remote Sensing Image Captioning
Remote sensing image (RSI) captioning aims to generate sentences to describe the content of RSIs. Generally, five sentences are used to describe the RSI in caption datasets.
Binqiang Wang +3 more
doaj +1 more source
In the age of social networks, the number of tweets sent by users has led to a sharp rise in public opinion. Public opinions are closely related to user stances. User stance detection has become an important task in the field of public opinion.
Peng Jia +5 more
doaj +1 more source
INTERACTIVE TOOL FOR VISUALIZATION OF TOPIC MODELS [PDF]
Digital data are all around us and occurs in various forms as videos, pictures or texts. Digital documents represent the vast majority of such data. It can be e-news, social media contributions and so on.
Miroslav SMATANA +3 more
doaj +1 more source
Analyzing LDA and NMF Topic Models for Urdu Tweets via Automatic Labeling
The understanding and analyzing of available content on Social media Platforms such as Twitter and Facebook, through various topic modeling methods is not supervised.
Zoya +3 more
doaj +1 more source
Data availability in articles published in Epidemiologia e Serviços de Saúde: a cross-sectional study, 2025 [PDF]
Objective: To assess the frequency and factors associated with data sharing in articles published in Epidemiologia e Serviços de Saúde. Methods: A cross-sectional study was conducted based on articles published in 2025 in the journal.
Taís Freire Galvão +1 more
doaj +4 more sources
Scientific Dataset Discovery via Topic-level Recommendation
Data intensive research requires the support of appropriate datasets. However, it is often time-consuming to discover usable datasets matching a specific research topic. We formulate the dataset discovery problem on an attributed heterogeneous graph, which is composed of paper-paper citation, paper-dataset citation, and also paper content.
Basmah Altaf +2 more
openaire +2 more sources
Cancer care is complex and exists within the broader healthcare system. The CanIMPACT team sought to enhance primary cancer care capacity and improve integration between primary and cancer specialist care, focusing on breast cancer.
Patti Ann Groome +7 more
doaj +1 more source
ClimaText: A Dataset for Climate Change Topic Detection
Accepted for the Tackling Climate Change with Machine Learning Workshop at NeurIPS ...
Varini, Francesco +3 more
openaire +3 more sources

