English
Related papers

Related papers: Vision Transformer for Multi-Domain Phase Retrieva…

200 papers

Fine-grained visual classification (FGVC) is a challenging computer vision problem, where the task is to automatically recognise objects from subordinate categories. One of its main difficulties is capturing the most discriminative…

Computer Vision and Pattern Recognition · Computer Science 2024-01-03 Dmitry Demidov , Muhammad Hamza Sharif , Aliakbar Abdurahimov , Hisham Cholakkal , Fahad Shahbaz Khan

Phase retrieval, a long-established challenge for recovering a complex-valued signal from its Fourier intensity measurements, has attracted significant interest because of its far-flung applications in optical imaging. To enhance accuracy,…

Signal Processing · Electrical Eng. & Systems 2023-05-16 Qiuliang Ye , Bingo Wing-Kuen Ling , Li-Wen Wang , Daniel Pak-Kong Lun

Vision Transformers (ViTs) achieve state-of-the-art segmentation accuracy but require large training datasets because each layer has unique parameters that must be learned independently. We present RD-ViT, a Recurrent-Depth Vision…

Computer Vision and Pattern Recognition · Computer Science 2026-05-06 Renjie He

In this manuscript we demonstrate a method to reconstruct the wavefront of focused beams from a measured diffraction pattern behind a diffracting mask in real-time. The phase problem is solved by means of a neural network, which is trained…

Image and Video Processing · Electrical Eng. & Systems 2021-04-07 Jonathon White , Sici Wang , Wilhelm Eschen , Jan Rothhardt

Vision Transformers (ViTs) have demonstrated strong potential in medical imaging; however, their high computational demands and tendency to overfit on small datasets limit their applicability in real-world clinical scenarios. In this paper,…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Aon Safdar , Mohamed Saadeldin

This paper considers the question of recovering the phase of an object from intensity-only measurements, a problem which naturally appears in X-ray crystallography and related disciplines. We study a physically realistic setup where one can…

Information Theory · Computer Science 2013-11-08 Emmanuel Candes , Xiaodong Li , Mahdi Soltanolkotabi

In the last decade, convolutional neural networks (ConvNets) have dominated and achieved state-of-the-art performances in a variety of medical imaging applications. However, the performances of ConvNets are still limited by lacking the…

Image and Video Processing · Electrical Eng. & Systems 2021-04-15 Junyu Chen , Yufan He , Eric C. Frey , Ye Li , Yong Du

While the implementation of single particle coherent diffraction imaging for non-crystalline particles is complicated by current limitations in photon flux, hit rate, and sample delivery a concept of many-particle coherent diffraction…

Disordered Systems and Neural Networks · Physics 2013-02-22 R. P. Kurta , R. Dronyak , M. Altarelli , E. Weckert , I. A. Vartanyants

Fraunhofer diffraction is a well-known phenomenon achieved with most wavelength even without lens. A single-shot intensity measurement of diffraction is generally considered inadequate to reconstruct the original light field, because the…

Image and Video Processing · Electrical Eng. & Systems 2019-12-04 An-Dong Xiong , Xiao-Peng Jin , Wen-Kai Yu , Qing Zhao

The rise of Deepfake technology to generate hyper-realistic manipulated images and videos poses a significant challenge to the public and relevant authorities. This study presents a robust Deepfake detection based on a modified Vision…

Computer Vision and Pattern Recognition · Computer Science 2025-08-28 Saksham Kumar , Rhythm Narang

With the advancement of deep learning technologies, specialized neural processing hardware such as Brain Processing Units (BPUs) have emerged as dedicated platforms for CNN acceleration, offering optimized INT8 computation capabilities for…

Computer Vision and Pattern Recognition · Computer Science 2026-02-09 Jinchi Tang , Yan Guo

We study a crucial yet often overlooked issue inherent to Vision Transformers (ViTs): feature maps of these models exhibit grid-like artifacts, which hurt the performance of ViTs in downstream dense prediction tasks such as semantic…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Jiawei Yang , Katie Z Luo , Jiefeng Li , Congyue Deng , Leonidas Guibas , Dilip Krishnan , Kilian Q Weinberger , Yonglong Tian , Yue Wang

Intra-frame inconsistency has been proved to be effective for the generalization of face forgery detection. However, learning to focus on these inconsistency requires extra pixel-level forged location annotations. Acquiring such annotations…

Computer Vision and Pattern Recognition · Computer Science 2022-10-25 Wanyi Zhuang , Qi Chu , Zhentao Tan , Qiankun Liu , Haojie Yuan , Changtao Miao , Zixiang Luo , Nenghai Yu

Rolling bearings are the most crucial components of rotating machinery. Identifying defective bearings in a timely manner may prevent the malfunction of an entire machinery system. The mechanical condition monitoring field has entered the…

Computer Vision and Pattern Recognition · Computer Science 2022-09-21 Abid Hasan Zim , Aeyan Ashraf , Aquib Iqbal , Asad Malik , Minoru Kuribayashi

Vision transformers (ViT) usually extract features via forwarding all the tokens in the self-attention layers from top to toe. In this paper, we introduce dynamic token-pass vision transformers (DoViT) for semantic segmentation, which can…

Computer Vision and Pattern Recognition · Computer Science 2023-08-25 Yuang Liu , Qiang Zhou , Jing Wang , Fan Wang , Jun Wang , Wei Zhang

Objective: Multi-shot interleaved echo planer imaging can obtain diffusion-weighted images (DWI) with high spatial resolution and low distortion, but suffers from ghost artifacts introduced by phase variations between shots. In this work,…

Signal Processing · Electrical Eng. & Systems 2022-12-09 Chen Qian , Zi Wang , Xinlin Zhang , Boxuan Shi , Boyu Jiang , Ran Tao , Jing Li , Yuwei Ge , Taishan Kang , Jianzhong Lin , Di Guo , Xiaobo Qu

We introduce a method to recover a continuous domain representation of a piecewise constant two-dimensional image from few low-pass Fourier samples. Assuming the edge set of the image is localized to the zero set of a trigonometric…

Computer Vision and Pattern Recognition · Computer Science 2016-04-25 Greg Ongie , Mathews Jacob

We present STCDiT, a video super-resolution framework built upon a pre-trained video diffusion model, aiming to restore structurally faithful and temporally stable videos from degraded inputs, even under complex camera motions. The main…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Junyang Chen , Jiangxin Dong , Long Sun , Yixin Yang , Jinshan Pan

Blind face restoration is a challenging task due to the unknown and complex degradation. Although face prior-based methods and reference-based methods have recently demonstrated high-quality results, the restored images tend to contain…

Computer Vision and Pattern Recognition · Computer Science 2024-03-01 Guojing Ge , Qi Song , Guibo Zhu , Yuting Zhang , Jinglu Chen , Miao Xin , Ming Tang , Jinqiao Wang

Representation learning with Vision Transformers (ViTs) has advanced rapidly, yet the utility of large-scale models in spatially sensitive tasks is hindered by spurious tokens. Prior efforts to mitigate this have been limited, often…

Computer Vision and Pattern Recognition · Computer Science 2026-05-20 Congpei Qiu , Zhaoyu Hu , Wei Ke , Zhuotao Tian , Yanhao Wu , Tong Zhang
‹ Prev 1 8 9 10 Next ›