Results 51 to 60 of about 11,102,357 (292)
A unified research data management framework for heterogeneous materials data is presented. The system integrates multimodal datasets using ontologies and knowledge graphs, enabling interoperability and FAIR (findable, accessible, interoperable, reusable) data principles. By linking data across scales and workflows, it supports reproducible, Artifitial
Doaa Mohamed +6 more
wiley +1 more source
DebateSum: A large-scale argument mining and summarization dataset
Prior work in Argument Mining frequently alludes to its potential applications in automatic debating systems. Despite this focus, almost no datasets or models exist which apply natural language processing techniques to problems found within competitive formal debate. To remedy this, we present the DebateSum dataset. DebateSum consists of 187,386 unique
Allen Roush, Arvind Balaji
openaire +3 more sources
Building machine‐readable vocabularies for materials science is slow, expert‐driven work. This study benchmarks 13 large language models on two of its first steps: finding candidate terms in engineering articles and deciding where they belong in a class hierarchy.
Thomas Bjarsch +3 more
wiley +1 more source
China’s iron ore deposits are generally characterized by substantial thickness and low grade, and this presents challenges for traditional mining methods to satisfy market demands economically and efficiently. To maintain a competitive edge in the market
Xiaocong YANG, Shenghua YIN
doaj +1 more source
The PRIMA Thesaurus for Materials Science and Engineering
The PRIMA Thesaurus is a structured vocabulary designed to improve how materials science data is described and shared. Developed with input from multiple experts, it enables clear documentation of research workflows, data exchange, and reuse across platforms.
Rossella Aversa +8 more
wiley +1 more source
Reproduction of stacking fault energy calculations from literature with a semi‐automated large language model‐assisted extraction procedure: extraction of simulation protocol, atomistic structures, computational parameters, and reported results, ontology alignment, knowledge graph construction and, finally, recomputation forvalidation.
Sepideh Baghaee Ravari +5 more
wiley +1 more source
The northern sand-proof belt is the key zone for preventing and controlling desertification and sandification in China, and it is the core area of the “Three Norths” project, which is characterized by infertile soil, high wind speed, drought and water ...
Quansheng LI +6 more
doaj +1 more source
Gold mining is a significant strategic sector for local, regional, and national economies. The rapid development of coexisting camp-type artisanal and small-scale gold mining (C-ASGM) and large-scale mining (LSM) accelerates the environmental and health ...
Satomi Kimijima +2 more
doaj +1 more source
A Practical Noise2Noise Denoising Pipeline for High‐Throughput Raman Spectroscopy
A lightweight and reproducible denoising pipeline for high‐throughput Raman spectroscopy is introduced, based on a 1D convolutional autoencoder trained with a Noise2Noise strategy. Using only repeated short‐exposure acquisitions, the method suppresses stochastic noise without reference spectra, enabling reliable spectral reconstruction while preserving
David Martin‐Calle +5 more
wiley +1 more source
Toolkit-based high-performance data mining of large data on MapReduce clusters
S.296-301The enormous growth of data in a variety of applications has increased the need for high performance data mining based on distributed environments. However, standard data mining toolkits per se do not allow the usage of computing clusters.
Wegener, Dennis +3 more
core +1 more source

