English
Related papers

Related papers: Steadiface: Real-Time Face-Centric Stabilization o…

200 papers

We present a novel appearance-based approach for pose estimation of a human hand using the point clouds provided by the low-cost Microsoft Kinect sensor. Both the free-hand case, in which the hand is isolated from the surrounding…

Computer Vision and Pattern Recognition · Computer Science 2016-04-08 Pasquale Coscia , Francesco A. N. Palmieri , Francesco Castaldo , Alberto Cavallo

Video stabilization is an in-camera processing commonly applied by modern acquisition devices. While significantly improving the visual quality of the resulting videos, it has been shown that such operation typically hinders the forensic…

Computer Vision and Pattern Recognition · Computer Science 2022-08-01 Andrea Montibeller , Cecilia Pasquini , Giulia Boato , Stefano Dell'Anna , Fernando Pérez-González

Action recognition in still images has seen major improvement in recent years due to advances in human pose estimation, object recognition and stronger feature representations produced by deep neural networks. However, there are still many…

Computer Vision and Pattern Recognition · Computer Science 2016-02-25 Amir Rosenfeld , Shimon Ullman

WiFi-based human pose estimation has emerged as a promising non-visual alternative approaches due to its pene-trability and privacy advantages. This paper presents VST-Pose, a novel deep learning framework for accurate and continuous pose…

Computer Vision and Pattern Recognition · Computer Science 2025-07-15 Xinyu Zhang , Zhonghao Ye , Jingwei Zhang , Xiang Tian , Zhisheng Liang , Shipeng Yu

Video summarization aims to generate a concise representation of a video, capturing its essential content and key moments while reducing its overall length. Although several methods employ attention mechanisms to handle long-term…

Computer Vision and Pattern Recognition · Computer Science 2024-05-22 Jaewon Son , Jaehun Park , Kwangsu Kim

In this paper, we propose a method, called GridFace, to reduce facial geometric variations and improve the recognition performance. Our method rectifies the face by local homography transformations, which are estimated by a face…

Computer Vision and Pattern Recognition · Computer Science 2018-08-21 Erjin Zhou , Zhimin Cao , Jian Sun

Accurate human localization is crucial for various applications, especially in the Metaverse era. Existing high precision solutions rely on expensive, tag-dependent hardware, while vision-based methods offer a cheaper, tag-free alternative.…

Computer Vision and Pattern Recognition · Computer Science 2025-02-19 Tianyi Zhang , Wengyu Zhang , Xulu Zhang , Jiaxin Wu , Xiao-Yong Wei , Jiannong Cao , Qing Li

This paper addresses the deep face recognition problem under an open-set protocol, where ideal face features are expected to have smaller maximal intra-class distance than minimal inter-class distance under a suitably chosen metric space.…

Computer Vision and Pattern Recognition · Computer Science 2022-03-17 Weiyang Liu , Yandong Wen , Bhiksha Raj , Rita Singh , Adrian Weller

Estimating pose of the head is an important preprocessing step in many pattern recognition and computer vision systems such as face recognition. Since the performance of the face recognition systems is greatly affected by the poses of the…

Computer Vision and Pattern Recognition · Computer Science 2012-05-15 Mohammad Tofighi , Hashem Kalbkhani , Mahrokh G. Shayesteh , Mehdi Ghasemzadeh

Image diffusion models are trained on independently sampled static images. While this is the bedrock task protocol in generative modeling, capturing the temporal world through the lens of static snapshots is information-deficient by design.…

Computer Vision and Pattern Recognition · Computer Science 2025-09-05 Juhun Lee , Simon S. Woo

To the best of our knowledge, we first present a live system that generates personalized photorealistic talking-head animation only driven by audio signals at over 30 fps. Our system contains three stages. The first stage is a deep neural…

Graphics · Computer Science 2021-09-27 Yuanxun Lu , Jinxiang Chai , Xun Cao

Video monocular depth estimation is essential for applications such as autonomous driving, AR/VR, and robotics. Recent transformer-based single-image monocular depth estimation models perform well on single images but struggle with depth…

Computer Vision and Pattern Recognition · Computer Science 2025-11-14 Sunghun Yang , Minhyeok Lee , Suhwan Cho , Jungho Lee , Sangyoun Lee

Face analysis is a core part of computer vision, in which remarkable progress has been observed in the past decades. Current methods achieve recognition and tracking with invariance to fundamental modes of variation such as illumination, 3D…

Computer Vision and Pattern Recognition · Computer Science 2018-03-12 Grigorios G. Chrysos , Paolo Favaro , Stefanos Zafeiriou

Current diffusion models for human image animation struggle to ensure identity (ID) consistency. This paper presents StableAnimator, the first end-to-end ID-preserving video diffusion framework, which synthesizes high-quality videos without…

Computer Vision and Pattern Recognition · Computer Science 2024-11-28 Shuyuan Tu , Zhen Xing , Xintong Han , Zhi-Qi Cheng , Qi Dai , Chong Luo , Zuxuan Wu

Tracking the pose of an object while it is being held and manipulated by a robot hand is difficult for vision-based methods due to significant occlusions. Prior works have explored using contact feedback and particle filters to localize…

Robotics · Computer Science 2020-11-09 Jacky Liang , Ankur Handa , Karl Van Wyk , Viktor Makoviychuk , Oliver Kroemer , Dieter Fox

As robotics progresses toward general manipulation, dexterous hands are becoming increasingly critical. However, proprioception in dexterous hands remains a bottleneck due to limitations in volume and generality. In this work, we present…

Robotics · Computer Science 2025-05-14 Junda Huang , Jianshu Zhou , Honghao Guo , Yunhui Liu

We propose an image-based, facial reenactment system that replaces the face of an actor in an existing target video with the face of a user from a source video, while preserving the original target performance. Our system is fully automatic…

Computer Vision and Pattern Recognition · Computer Science 2016-02-09 Pablo Garrido , Levi Valgaerts , Ole Rehmsen , Thorsten Thormaehlen , Patrick Perez , Christian Theobalt

Whole-body pose estimation localizes the human body, hand, face, and foot keypoints in an image. This task is challenging due to multi-scale body parts, fine-grained localization for low-resolution regions, and data scarcity. Meanwhile,…

Computer Vision and Pattern Recognition · Computer Science 2023-08-28 Zhendong Yang , Ailing Zeng , Chun Yuan , Yu Li

Face Super-Resolution (SR) is a subfield of the SR domain that specifically targets the reconstruction of face images. The main challenge of face SR is to restore essential facial features without distortion. We propose a novel face SR…

Computer Vision and Pattern Recognition · Computer Science 2019-08-23 Deokyun Kim , Minseon Kim , Gihyun Kwon , Dae-Shik Kim

Talking face generation aims to create realistic videos with accurate lip synchronization and high visual quality, using given audio and reference video while preserving identity and visual characteristics. In this paper, we start by…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Dogucan Yaman , Fevziye Irem Eyiokur , Leonard Bärmann , Hazim Kemal Ekenel , Alexander Waibel
‹ Prev 1 8 9 10 Next ›