Results 61 to 70 of about 25,150 (262)

Evaluation of vision transformers for the detection of fullness of garbage bins for efficient waste management

open access: yesFrontiers in Artificial Intelligence
Efficient waste management is crucial for urban environments to maintain cleanliness, reduce environmental impact, and optimize resource allocation. Traditional waste collection systems often rely on scheduled pickups or manual inspections, leading to ...
Parakram Singh Tanwer   +4 more
doaj   +1 more source

Transformer-based progressive residual network for single image dehazing

open access: yesFrontiers in Neurorobotics, 2022
IntroductionThe seriously degraded fogging image affects the further visual tasks. How to obtain a fog-free image is not only challenging, but also important in computer vision.
Zhe Yang   +4 more
doaj   +1 more source

Denoising Vision Transformers

open access: yes
Accepted to ECCV2024. Project website: https://jiawei-yang.github.io/DenoisingViT/
Jiawei Yang 0002   +8 more
openaire   +2 more sources

What Do Large Language Models Know About Materials?

open access: yesAdvanced Engineering Materials, EarlyView.
If large language models (LLMs) are to be used inside the material discovery and engineering process, they must be benchmarked for the accurateness of intrinsic material knowledge. The current work introduces 1) a reasoning process through the processing–structure–property–performance chain and 2) a tool for benchmarking knowledge of LLMs concerning ...
Adrian Ehrenhofer   +2 more
wiley   +1 more source

NFDI MatWerk Ontology (MWO): A BFO‐Compliant Ontology for Research Data Management in Materials Science and Engineering

open access: yesAdvanced Engineering Materials, EarlyView.
This article presents the NFDI‐MatWerk Ontology (MWO), a Basic Formal Ontology‐based framework for interoperable research data management in materials science and engineering (MSE). Covering consortium structures, research data management resources, services, and instruments, MWO enables semantic integration, Findable, Accessible, Interoperable, and ...
Hossein Beygi Nasrabadi   +4 more
wiley   +1 more source

Discrete Wavelet Transform Meets Transformer: Unleashing the Full Potential of the Transformer for Visual Recognition

open access: yesIEEE Access, 2023
Traditionally, the success of the Transformer has been attributed to its token mixer, particularly the self-attention mechanism. However, recent studies suggest that replacing such attention-based token mixer with alternative techniques can yield ...
Dongwook Yang, Seung-Woo Seo
doaj   +1 more source

Grafting Vision Transformers

open access: yes2024 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)
Vision Transformers (ViTs) have recently become the state-of-the-art across many computer vision tasks. In contrast to convolutional networks (CNNs), ViTs enable global information sharing even within shallow layers of a network, i.e., among high-resolution features. However, this perk was later overlooked with the success of pyramid architectures such
Jongwoo Park 0003   +5 more
openaire   +2 more sources

Vision Language Transformers: A Survey

open access: yesCoRR, 2023
Vision language tasks, such as answering questions about or generating captions that describe an image, are difficult tasks for computers to perform. A relatively recent body of research has adapted the pretrained transformer architecture introduced in \citet{vaswani2017attention} to vision language modeling.
Clayton Fields, Casey Kennington
openaire   +2 more sources

The PRIMA Thesaurus for Materials Science and Engineering

open access: yesAdvanced Engineering Materials, EarlyView.
The PRIMA Thesaurus is a structured vocabulary designed to improve how materials science data is described and shared. Developed with input from multiple experts, it enables clear documentation of research workflows, data exchange, and reuse across platforms.
Rossella Aversa   +8 more
wiley   +1 more source

SE‐Swin: An improved Swin‐Transfomer network of self‐ensemble feature extraction framework for image retrieval

open access: yesIET Image Processing
The Swin‐Transformer is a variant of the Vision Transformer, which constructs a hierarchical Transformer that computes representations with shifted windows and window multi‐head self‐attention.
Yixuan Xu   +3 more
doaj   +1 more source

Home - About - Disclaimer - Privacy