English
Related papers

Related papers: Self-Supervised Robustifying Guidance for Monocula…

200 papers

In this paper, a novel architecture for speaker recognition is proposed by cascading speech enhancement and speaker processing. Its aim is to improve speaker recognition performance when speech signals are corrupted by noise. Instead of…

Computation and Language · Computer Science 2020-05-25 Yanpei Shi , Qiang Huang , Thomas Hain

3D human pose and shape recovery from a monocular RGB image is a challenging task. Existing learning based methods highly depend on weak supervision signals, e.g. 2D and 3D joint location, due to the lack of in-the-wild paired 3D…

Computer Vision and Pattern Recognition · Computer Science 2021-12-28 Zhiwei Liu , Xiangyu Zhu , Lu Yang , Xiang Yan , Ming Tang , Zhen Lei , Guibo Zhu , Xuetao Feng , Yan Wang , Jinqiao Wang

Self-supervised learning can be used for mitigating the greedy needs of Vision Transformer networks for very large fully-annotated datasets. Different classes of self-supervised learning offer representations with either good contextual…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Spyros Gidaris , Andrei Bursuc , Oriane Simeoni , Antonin Vobecky , Nikos Komodakis , Matthieu Cord , Patrick Pérez

Generative reconstruction methods compute the 3D configuration (such as pose and/or geometry) of a shape by optimizing the overlap of the projected 3D shape model with images. Proper handling of occlusions is a big challenge, since the…

Computer Vision and Pattern Recognition · Computer Science 2016-02-12 Helge Rhodin , Nadia Robertini , Christian Richardt , Hans-Peter Seidel , Christian Theobalt

Fundus image classification is crucial in the computer aided diagnosis tasks, but label noise significantly impairs the performance of deep neural networks. To address this challenge, we propose a robust framework, Self-Supervised…

Computer Vision and Pattern Recognition · Computer Science 2024-10-25 Mengwen Ye , Yingzi Huangfu , You Li , Zekuan Yu

Deformable surface tracking from monocular images is well-known to be under-constrained. Occlusions often make the task even more challenging, and can result in failure if the surface is not sufficiently textured. In this work, we…

Computer Vision and Pattern Recognition · Computer Science 2015-09-28 Dat Tien Ngo , Sanghuyk Park , Anne Jorstad , Alberto Crivellaro , Chang Yoo , Pascal Fua

Despite recent advances in large-scale text-to-image generative models, manipulating real images with these models remains a challenging problem. The main limitations of existing editing methods are that they either fail to perform with…

Computer Vision and Pattern Recognition · Computer Science 2024-09-26 Vadim Titov , Madina Khalmatova , Alexandra Ivanova , Dmitry Vetrov , Aibek Alanov

Current state-of-the-art methods cast monocular 3D human pose estimation as a learning problem by training neural networks on large data sets of images and corresponding skeleton poses. In contrast, we propose an approach that can exploit…

Computer Vision and Pattern Recognition · Computer Science 2020-10-14 Simon Jenni , Paolo Favaro

Synthesizing normal-light novel views from low-light multiview images is an important yet challenging task, given the low visibility and high ISO noise present in the input images. Existing low-light enhancement methods often struggle to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-17 Ze Li , Feng Zhang , Xiatian Zhu , Meng Zhang , Yanghong Zhou , P. Y. Mok

In this work, we propose a method that combines unsupervised deep learning predictions for optical flow and monocular disparity with a model based optimization procedure for instantaneous camera pose. Given the flow and disparity…

Computer Vision and Pattern Recognition · Computer Science 2019-02-14 Alex Zihao Zhu , Wenxin Liu , Ziyun Wang , Vijay Kumar , Kostas Daniilidis

We present an approach to estimating camera rotation in crowded, real-world scenes from handheld monocular video. While camera rotation estimation is a well-studied problem, no previous methods exhibit both high accuracy and acceptable…

Computer Vision and Pattern Recognition · Computer Science 2023-09-18 Fabien Delattre , David Dirnfeld , Phat Nguyen , Stephen Scarano , Michael J. Jones , Pedro Miraldo , Erik Learned-Miller

The discriminability of feature representation is the key to open-set face recognition. Previous methods rely on the learnable weights of the classification layer that represent the identities. However, the evaluation process learns no…

Computer Vision and Pattern Recognition · Computer Science 2023-04-25 Youzhe Song , Feng Wang

Audio-driven IDentity-specific Talking Head Generation (ID-specific THG) has shown increasing promise for applications in filmmaking and virtual reality. Existing approaches are generally constructed as end-to-end paradigms, and have…

Computer Vision and Pattern Recognition · Computer Science 2025-06-10 Jian Yang , Xukun Wang , Wentao Wang , Guoming Li , Qihang Fang , Ruihong Yuan , Tianyang Wang , Xiaomei Zhang , Yeying Jin , Zhaoxin Fan

Accurate depth estimation under out-of-distribution (OoD) scenarios, such as adverse weather conditions, sensor failure, and noise contamination, is desirable for safety-critical applications. Existing depth estimation systems, however,…

Depth-guided 3D reconstruction has gained popularity as a fast alternative to optimization-heavy approaches, yet existing methods still suffer from scale drift, multi-view inconsistencies, and the need for substantial refinement to achieve…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Kang Han , Wei Xiang , Lu Yu , Mathew Wyatt , Gaowen Liu , Ramana Rao Kompella

3D human pose estimation using monocular images is an important yet challenging task. Existing 3D pose detection methods exhibit excellent performance under normal conditions however their performance may degrade due to occlusion. Recently…

Computer Vision and Pattern Recognition · Computer Science 2022-03-09 Mehwish Ghafoor , Arif Mahmood

Three-dimensional face dense alignment and reconstruction in the wild is a challenging problem as partial facial information is commonly missing in occluded and large pose face images. Large head pose variations also increase the solution…

Computer Vision and Pattern Recognition · Computer Science 2021-07-07 Zeyu Ruan , Changqing Zou , Longhai Wu , Gangshan Wu , Limin Wang

Unsupervised Camouflaged Object Detection (UCOD) remains a challenging task due to the high intrinsic similarity between target objects and their surroundings, as well as the reliance on noisy pseudo-labels that hinder fine-grained texture…

Computer Vision and Pattern Recognition · Computer Science 2026-03-13 Shuo Jiang , Gaojia Zhang , Min Tan , Yufei Yin , Gang Pan

Person re-identification is vital for monitoring and tracking crowd movement to enhance public security. However, re-identification in the presence of occlusion substantially reduces the performance of existing systems and is a challenging…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Prathistith Raj Medi , Ghanta Sai Krishna , Praneeth Nemani , Satyanarayana Vollala , Santosh Kumar

There are many facts affecting human face recognition, such as pose, occlusion, illumination, age, etc. First and foremost are large pose and occlusion problems, which can even result in more than 10% performance degradation. Pose-invariant…

Computer Vision and Pattern Recognition · Computer Science 2019-02-27 Qingyan Duan , Lei Zhang
‹ Prev 1 8 9 10 Next ›