English
Related papers

Related papers: Vision Transformer for Multi-Domain Phase Retrieva…

200 papers

We consider the imaging problem of the reconstruction of a three-dimensional object via optical diffraction tomography under the assumptions of the Born approximation. Our focus lies in the situation that a rigid object performs an…

Numerical Analysis · Mathematics 2024-07-11 Robert Beinert , Michael Quellmalz

Partial scan is a common approach to accelerate Magnetic Resonance Imaging (MRI) data acquisition in both 2D and 3D settings. However, accurately reconstructing images from partial scan data (i.e., incomplete k-space matrices) remains…

Image and Video Processing · Electrical Eng. & Systems 2023-06-06 Xiaohan Liu , Yanwei Pang , Xuebin Sun , Yiming Liu , Yonghong Hou , Zhenchang Wang , Xuelong Li

The Vision Transformer (ViT) architecture has established its place in computer vision literature, however, training ViTs for RGB-D object recognition remains an understudied topic, viewed in recent literature only through the lens of…

Computer Vision and Pattern Recognition · Computer Science 2023-03-08 Georgios Tziafas , Hamidreza Kasaei

Cross-subject motor imagery (CS-MI) classification in brain-computer interfaces (BCIs) is a challenging task due to the significant variability in Electroencephalography (EEG) patterns across different individuals. This variability often…

Machine Learning · Computer Science 2025-07-04 Ahmed G. Habashi , Ahmed M. Azab , Seif Eldawlatly , Gamal M. Aly

Computer vision methods that explicitly detect object parts and reason on them are a step towards inherently interpretable models. Existing approaches that perform part discovery driven by a fine-grained classification task make very…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Ananthu Aniraj , Cassio F. Dantas , Dino Ienco , Diego Marcos

Low-dose computed tomography (LDCT) reduces radiation exposure but suffers from image artifacts and loss of detail due to quantum and electronic noise, potentially impacting diagnostic accuracy. Transformer combined with diffusion models…

Image and Video Processing · Electrical Eng. & Systems 2025-07-01 Qiqing Liu , Guoquan Wei , Zekun Zhou , Yiyang Wen , Liu Shi , Qiegen Liu

Vision Transformers (ViTs) have achieved overwhelming success, yet they suffer from vulnerable resolution scalability, i.e., the performance drops drastically when presented with input resolutions that are unseen during training. We…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Rui Tian , Zuxuan Wu , Qi Dai , Han Hu , Yu Qiao , Yu-Gang Jiang

Diffusion models have demonstrated their utility as learned priors for solving various inverse problems. However, most existing approaches are limited to linear inverse problems. This paper exploits the efficient and unsupervised posterior…

Image and Video Processing · Electrical Eng. & Systems 2025-01-07 Mehmet Onurcan Kaya , Figen S. Oktem

Twinning is a common crystallographic phenomenon, which usually occurs in crystals during symmetry-lowering phase transition. Once formed, twin domains play an important role in defining physical properties: for example, twin domains…

Materials Science · Physics 2021-10-28 Semen Gorfman , David Spirito , Guanjie Zhang , Carsten Detlefs , Nan Zhang

Active research is currently underway to enhance the efficiency of vision transformers (ViTs). Most studies have focused solely on effective token mixers, overlooking the potential relationship with normalization. To boost diverse feature…

Computer Vision and Pattern Recognition · Computer Science 2024-12-02 Jongseong Bae , Susang Kim , Minsu Cho , Ha Young Kim

Vision Transformer (ViT) has achieved remarkable success due to its large-scale pretraining on general domains, but it still faces challenges when applying it to downstream distant domains that have only scarce training data, which gives…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Shuai Yi , Yixiong Zou , Yuhua Li , Ruixuan Li

Coherent diffraction imaging (CDI) on Bragg reflections is a promising technique for the study of three-dimensional (3D) composition and strain fields in nanostructures, which can be recovered directly from the coherent diffraction data…

The Vision Transformer (ViT) excels in accuracy when handling high-resolution images, yet it confronts the challenge of significant spatial redundancy, leading to increased computational and memory requirements. To address this, we present…

Computer Vision and Pattern Recognition · Computer Science 2024-02-02 Youbing Hu , Yun Cheng , Anqi Lu , Zhiqiang Cao , Dawei Wei , Jie Liu , Zhijun Li

Lattice defects play a key role in determining the properties of crystalline materials. Probing the 3D lattice strains that govern their interactions remains a challenge. Bragg Coherent Diffraction Imaging (BCDI) allows strain to be…

High-contrast imaging relies on advanced coronagraphs and adaptive optics (AO) to attenuate the starlight. However, residual aberrations, especially non-common path aberrations between the AO channel and the coronagraph channel, limit the…

Instrumentation and Methods for Astrophysics · Physics 2025-11-11 Axel Potier , Raphaël Galicher , Pierre Baudoz , Johan Mazoyer , Zahed Wahhaj , Ruben Tandon , Jonas G. Kühn , Laura Perez , Gael Chauvin

This paper develops a novel framework for phase retrieval, a problem which arises in X-ray crystallography, diffraction imaging, astronomical imaging and many other applications. Our approach combines multiple structured illuminations…

Information Theory · Computer Science 2011-09-21 Emmanuel J. Candes , Yonina Eldar , Thomas Strohmer , Vlad Voroninski

In recent years, the rapid advancement of deepfake technology has revolutionized content creation, lowering forgery costs while elevating quality. However, this progress brings forth pressing concerns such as infringements on individual…

Computer Vision and Pattern Recognition · Computer Science 2024-05-15 Zhikan Wang , Zhongyao Cheng , Jiajie Xiong , Xun Xu , Tianrui Li , Bharadwaj Veeravalli , Xulei Yang

The rapid advancement of generative models has led to a growing prevalence of highly realistic AI-generated images, posing significant challenges for digital forensics and content authentication. Conventional detection methods mainly rely…

Computer Vision and Pattern Recognition · Computer Science 2025-08-26 Dabbrata Das , Mahshar Yahan , Md Tareq Zaman , Md Rishadul Bayesh

Studies have proven that domain bias and label bias exist in different Facial Expression Recognition (FER) datasets, making it hard to improve the performance of a specific dataset by adding other datasets. For the FER bias issue, recent…

Computer Vision and Pattern Recognition · Computer Science 2022-11-15 Shuyi Mao , Xinpeng Li , Qingyang Wu , Xiaojiang Peng

The hybrid deep models of Vision Transformer (ViT) and Convolution Neural Network (CNN) have emerged as a powerful class of backbones for vision tasks. Scaling up the input resolution of such hybrid backbones naturally strengthes model…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Ting Yao , Yehao Li , Yingwei Pan , Tao Mei
‹ Prev 1 3 4 5 6 7 10 Next ›