English
Related papers

Related papers: FAFA: Frequency-Aware Flow-Aided Self-Supervision …

200 papers

Object detection models typically perform well on images captured in controlled environments with stable lighting, water clarity, and viewpoint, but their performance degrades substantially in real-world underwater settings characterized by…

Computer Vision and Pattern Recognition · Computer Science 2026-04-24 Eleanor Wiesler , Trace Baxley

One of the core activities of an active observer involves moving to secure a "better" view of the scene, where the definition of "better" is task-dependent. This paper focuses on the task of human pose estimation from videos capturing a…

Robotics · Computer Science 2024-07-03 Jingxi Chen , Botao He , Chahat Deep Singh , Cornelia Fermuller , Yiannis Aloimonos

By using unsupervised domain adaptation (UDA), knowledge can be transferred from a label-rich source domain to a target domain that contains relevant information but lacks labels. Many existing UDA algorithms suffer from directly using raw…

Computer Vision and Pattern Recognition · Computer Science 2024-03-22 Le Luo , Bingrong Xu , Qingyong Zhang , Cheng Lian , Jie Luo

In this work we propose a one-class self-supervised method for anomaly segmentation in images that benefits both from a modern machine learning approach and a more classic statistical detection theory. The method consists of four phases.…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Matías Tailanian , Álvaro Pardo , Pablo Musé

Multimodal image fusion effectively aggregates information from diverse modalities, with fused images playing a crucial role in vision systems. However, existing methods often neglect frequency-domain feature exploration and interactive…

Computer Vision and Pattern Recognition · Computer Science 2025-06-05 Tianpei Zhang , Jufeng Zhao , Yiming Zhu , Guangmang Cui

Human pose estimation (HPE) is a central part of understanding the visual narration and body movements of characters depicted in artwork collections, such as Greek vase paintings. Unfortunately, existing HPE methods do not generalise well…

Computer Vision and Pattern Recognition · Computer Science 2024-02-27 Prathmesh Madhu , Angel Villar-Corrales , Ronak Kosti , Torsten Bendschus , Corinna Reinhardt , Peter Bell , Andreas Maier , Vincent Christlein

Underwater visuals undergo various complex degradations, inevitably influencing the efficiency of underwater vision tasks. Recently, diffusion models were employed to underwater image enhancement (UIE) tasks, and gained SOTA performance.…

Computer Vision and Pattern Recognition · Computer Science 2026-02-13 Chen Zhao , Chenyu Dong , Weiling Cai , Yueyue Wang

This work reviews the problem of object detection in underwater environments. We analyse and quantify the shortcomings of conventional state-of-the-art (SOTA) algorithms in the computer vision community when applied to this challenging…

Computer Vision and Pattern Recognition · Computer Science 2022-05-24 Andre Jesus , Claudio Zito , Claudio Tortorici , Eloy Roura , Giulia De Masi

We present an algorithm, Fourier Activity Recognition (FAR), for UAV video activity recognition. Our formulation uses a novel Fourier object disentanglement method to innately separate out the human agent (which is typically small) from the…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Divya Kothandaraman , Tianrui Guan , Xijun Wang , Sean Hu , Ming Lin , Dinesh Manocha

The FlowNet demonstrated that optical flow estimation can be cast as a learning problem. However, the state of the art with regard to the quality of the flow has still been defined by traditional methods. Particularly on small displacements…

Computer Vision and Pattern Recognition · Computer Science 2016-12-07 Eddy Ilg , Nikolaus Mayer , Tonmoy Saikia , Margret Keuper , Alexey Dosovitskiy , Thomas Brox

We propose a learning-based depth from focus/defocus (DFF), which takes a focal stack as input for estimating scene depth. Defocus blur is a useful cue for depth estimation. However, the size of the blur depends on not only scene depth but…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Yuki Fujimura , Masaaki Iiyama , Takuya Funatomi , Yasuhiro Mukaigawa

Millimeter-Wave (mmWave) radar can enable high-resolution human pose estimation with low cost and computational requirements. However, mmWave data point cloud, the primary input to processing algorithms, is highly sparse and carries…

Image and Video Processing · Electrical Eng. & Systems 2022-05-03 Sizhe An , Umit Y. Ogras

Human Pose Estimation (HPE) is increasingly important for applications like virtual reality and motion analysis, yet current methods struggle with balancing accuracy, computational efficiency, and reliable uncertainty quantification (UQ).…

Computer Vision and Pattern Recognition · Computer Science 2026-01-30 Shipeng Liu , Ziliang Xiong , Bastian Wandt , Per-Erik Forssén

Attention calculation is extremely time-consuming for long-sequence inference tasks, such as text or image/video generation, in large models. To accelerate this process, we developed a low-precision, mathematically-equivalent algorithm…

This paper addresses the problem of estimating the 3-DoF camera pose for a ground-level image with respect to a satellite image that encompasses the local surroundings. We propose a novel end-to-end approach that leverages the learning of…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Zhenbo Song , Xianghui Ze , Jianfeng Lu , Yujiao Shi

Estimating the 3D motion of points in a scene, known as scene flow, is a core problem in computer vision. Traditional learning-based methods designed to learn end-to-end 3D flow often suffer from poor generalization. Here we present a…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Yair Kittenplon , Yonina C. Eldar , Dan Raviv

Depth estimation aims to predict dense depth maps. In autonomous driving scenes, sparsity of annotations makes the task challenging. Supervised models produce concave objects due to insufficient structural information. They overfit to valid…

Computer Vision and Pattern Recognition · Computer Science 2023-08-07 Jiaqi Li , Yiran Wang , Zihao Huang , Jinghong Zheng , Ke Xian , Zhiguo Cao , Jianming Zhang

Recent methods for long-tailed instance segmentation still struggle on rare object classes with few training data. We propose a simple yet effective method, Feature Augmentation and Sampling Adaptation (FASA), that addresses the data…

Computer Vision and Pattern Recognition · Computer Science 2021-10-01 Yuhang Zang , Chen Huang , Chen Change Loy

3D hand-object pose estimation is an important issue to understand the interaction between human and environment. Current hand-object pose estimation methods require detailed 3D labels, which are expensive and labor-intensive. To tackle the…

Computer Vision and Pattern Recognition · Computer Science 2021-07-19 Zida Cheng , Siheng Chen , Ya Zhang

Fast fluid antenna multiple access (FAMA) is an idea that promises to overcome severe interference in massive access scenarios by reconfiguring the antenna's position at the receiver side on a symbol-by-symbol basis, without the need of…

Signal Processing · Electrical Eng. & Systems 2026-05-25 Noor Waqar , Kai-Kit Wong , Chan-Byoung Chae , Ross Murch