Results 41 to 50 of about 1,373,967 (287)
Survey of Vision Transformers(ViT) [PDF]
The Vision Transformer(ViT),an application of the Transformer architecture with an encoder-decoder structure,has garnered remarkable success in the field of computer vision.Over the past few years,research centered around ViT has witnessed a prolific ...
LI Yujie, MA Zihang, WANG Yifu, WANG Xinghe, TAN Benying
doaj +1 more source
Measurements With A Quantum Vision Transformer: A Naive Approach [PDF]
In mainstream machine learning, transformers are gaining widespread usage. As Vision Transformers rise in popularity in computer vision, they now aim to tackle a wide variety of machine learning applications.
Pasquali Dominic +2 more
doaj +1 more source
Vision Transformers in Image Restoration: A Survey
The Vision Transformer (ViT) architecture has been remarkably successful in image restoration. For a while, Convolutional Neural Networks (CNN) predominated in most computer vision tasks.
Anas M. Ali +5 more
doaj +1 more source
QuadTree Attention for Vision Transformers
Transformers have been successful in many vision tasks, thanks to their capability of capturing long-range dependency. However, their quadratic computational complexity poses a major obstacle for applying them to vision tasks requiring dense predictions, such as object detection, feature matching, stereo, etc.
Tang, Shitao +3 more
openaire +4 more sources
3D-Vision-Transformer Stacking Ensemble for Assessing Prostate Cancer Aggressiveness from T2w Images
Vision transformers represent the cutting-edge topic in computer vision and are usually employed on two-dimensional data following a transfer learning approach.
Eva Pachetti, Sara Colantonio
doaj +1 more source
Variable-Rate Deep Image Compression With Vision Transformers
Recently, vision transformers have been applied in many computer vision problems due to its long-range learning ability. However, it has not been throughly explored in image compression.
Binglin Li, Jie Liang, Jingning Han
doaj +1 more source
Semi-supervised Vision Transformers
We study the training of Vision Transformers for semi-supervised image classification. Transformers have recently demonstrated impressive performance on a multitude of supervised learning tasks. Surprisingly, we show Vision Transformers perform significantly worse than Convolutional Neural Networks when only a small set of labeled data is available ...
Zejia Weng +4 more
openaire +3 more sources
Bragg grating based integrated photonic Hilbert transformers
Planar Bragg grating based photonic Hilbert transformers are experimentally demonstrated in this work. Planar Bragg gratings are utilized to implement a general Hilbert transform, a fractional order Hilbert transform and terahertz bandwidth Hilbert ...
Smith, P.G.R. +3 more
core +2 more sources
EnViTSA: Ensemble of Vision Transformer with SpecAugment for Acoustic Event Classification
Recent successes in deep learning have inspired researchers to apply deep neural networks to Acoustic Event Classification (AEC). While deep learning methods can train effective AEC models, they are susceptible to overfitting due to the models’ high ...
Kian Ming Lim +3 more
doaj +1 more source
Vision Transformers (ViTs) have recently become the state-of-the-art across many computer vision tasks. In contrast to convolutional networks (CNNs), ViTs enable global information sharing even within shallow layers of a network, i.e., among high-resolution features. However, this perk was later overlooked with the success of pyramid architectures such
Jongwoo Park 0003 +5 more
openaire +2 more sources

