Results 71 to 80 of about 25,150 (262)

A Hierarchical ViT With Dynamic Window Shift Unit and Curriculum Learning for Remote Sensing Image Scene Classification

open access: yesIEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing
The remote sensing image (RSI) scene classification is currently a popular research topic among many remote sensing tasks. However, RSI scene classification still faces challenges such as complex multiscale key features concentrated in different local ...
Yi Liu   +5 more
doaj   +1 more source

Ontology‐Aligned Structuring and Reuse of Multimodal Materials Data and Workflows Toward Automatic Reproduction

open access: yesAdvanced Engineering Materials, EarlyView.
Reproduction of stacking fault energy calculations from literature with a semi‐automated large language model‐assisted extraction procedure: extraction of simulation protocol, atomistic structures, computational parameters, and reported results, ontology alignment, knowledge graph construction and, finally, recomputation forvalidation.
Sepideh Baghaee Ravari   +5 more
wiley   +1 more source

Classification of Lung Diseases in X-Ray Images Using Transformer-Based Deep Learning Models

open access: yesJurnal Nasional Pendidikan Teknik Informatika (JANAPATI)
This research evaluates the performance of two Transformer models, the Vision Transformer (ViT) and Swin Transformer, in the analysis of thoracic X-ray images.
Nyoman Sarasuartha Mahajaya   +2 more
doaj   +1 more source

Machine Learning‐Supported Analysis for Predicting and Visualizing Nonlinear Relationships Between Material Properties in Electroplated Chromium Layers

open access: yesAdvanced Engineering Materials, EarlyView.
This study applies machine learning regression to predict chromium layer thickness in decorative trivalent chromium electroplating, using 441 experiments from laboratory‐scale (1L) and pilot‐scale (14L) setups. Tree‐based models, particularly CatBoost, outperformed linear regression by capturing nonlinear parameter interactions (R2$R^2$ up to 0.77 ...
Christoph Baumer   +4 more
wiley   +1 more source

A Sensorimotor Vision Transformer

open access: yesCoRR
This paper presents the Sensorimotor Transformer (SMT), a vision model inspired by human saccadic eye movements that prioritize high-saliency regions in visual input to enhance computational efficiency and reduce memory consumption. Unlike traditional models that process all image patches uniformly, SMT identifies and selects the most salient patches ...
Konrad Gadzicki   +2 more
openaire   +2 more sources

DigiChrom: A Domain Ontology for Semantic Representation of Trivalent Chromium Platings and Its Large Language Model‐Based Alignment With Multiple Mid‐Level Ontologies

open access: yesAdvanced Engineering Materials, EarlyView.
Digitalizing electroplating requires both domain knowledge and interoperability. This work introduces PlatOn, a domain ontology for trivalent chromium plating and coating characterization, and a hybrid pipeline that aligns it to a mid‐level reference ontology by combining eight similarity metrics with language model reasoning. Expert‐validated mappings
Janik Harter   +10 more
wiley   +1 more source

Semantic Modeling in Materials Science and Engineering With Platform MaterialDigital Core Ontology 3.0

open access: yesAdvanced Engineering Materials, EarlyView.
The community‐driven Platform MaterialDigital Core Ontology (PMDco) 3.0 is introduced as a Basic Formal Ontology‐aligned semantic backbone for the processing–structure–properties paradigm in Materials Science and Engineering. Modular engineering, automated releases, and validation workflows are highlighted and key semantic patterns for materials ...
Markus Schilling   +15 more
wiley   +1 more source

Hybrid model integrating LeViT transformer and distillation techniques for pattern detection and dance classification

open access: yesScientific Reports
Artificial intelligence (AI) has become an integral part of modern life, extending its impact into the preservation of cultural heritage. This study applies state-of-the-art vision transformer models for the classification of traditional Chinese dance ...
Yanyan Wang
doaj   +1 more source

FSwin Transformer: Feature-Space Window Attention Vision Transformer for Image Classification

open access: yesIEEE Access
The vision transformer (ViT) with global self-attention exhibits quadratic computational complexity that depends on the image size. To address this issue, window-based self-attention ViT limits attention area to a specific window, thereby mitigating the ...
Dayeon Yoo, Jeesu Kim, Jinwoo Yoo
doaj   +1 more source

Depth Perception Using Various Vision Transformer [PDF]

open access: yesEPJ Web of Conferences
Proper depth perception is one of the key requirements of three-dimensional understanding of scenes in the context of self-driving. The discussed manuscript defines a re-architecturing of VoxelNet with a dual attention paradigm (inspired by Vision ...
Kukreja Swetta   +4 more
doaj   +1 more source

Home - About - Disclaimer - Privacy