English
Related papers

Related papers: Reshoot-Anything: A Self-Supervised Model for In-t…

200 papers

In this paper, we propose a novel learning approach for feed-forward one-shot 4D head avatar synthesis. Different from existing methods that often learn from reconstructing monocular videos guided by 3DMM, we employ pseudo multi-view videos…

Computer Vision and Pattern Recognition · Computer Science 2024-07-12 Yu Deng , Duomin Wang , Baoyuan Wang

Autoregressive video diffusion models hold promise for world simulation but are vulnerable to exposure bias arising from the train-test mismatch. While recent works address this via post-training, they typically rely on a bidirectional…

Computer Vision and Pattern Recognition · Computer Science 2025-12-18 Yuwei Guo , Ceyuan Yang , Hao He , Yang Zhao , Meng Wei , Zhenheng Yang , Weilin Huang , Dahua Lin

A key question in the problem of 3D reconstruction is how to train a machine or a robot to model 3D objects. Many tasks like navigation in real-time systems such as autonomous vehicles directly depend on this problem. These systems usually…

Computer Vision and Pattern Recognition · Computer Science 2022-09-22 AmirHossein Zamani , Amir G. Aghdam , Kamran Ghaffari T

We propose an image-based, facial reenactment system that replaces the face of an actor in an existing target video with the face of a user from a source video, while preserving the original target performance. Our system is fully automatic…

Computer Vision and Pattern Recognition · Computer Science 2016-02-09 Pablo Garrido , Levi Valgaerts , Ole Rehmsen , Thorsten Thormaehlen , Patrick Perez , Christian Theobalt

Untrimmed videos on social media or those captured by robots and surveillance cameras are of varied aspect ratios. However, 3D CNNs usually require as input a square-shaped video, whose spatial dimension is smaller than the original.…

Computer Vision and Pattern Recognition · Computer Science 2021-11-23 Prithwish Jana , Swarnabja Bhaumik , Partha Pratim Mohanta

We explore novel-view synthesis for dynamic scenes from monocular videos. Prior approaches rely on costly test-time optimization of 4D representations or do not preserve scene geometry when trained in a feed-forward manner. Our approach is…

Computer Vision and Pattern Recognition · Computer Science 2026-01-14 Kaihua Chen , Tarasha Khurana , Deva Ramanan

Recovering 4D world from monocular video is a crucial yet challenging task. Conventional methods usually rely on the assumptions of multi-view videos, known camera parameters, or static scenes. In this paper, we relax all these constraints…

Computer Vision and Pattern Recognition · Computer Science 2025-01-03 Shizun Wang , Xingyi Yang , Qiuhong Shen , Zhenxiang Jiang , Xinchao Wang

3D photography renders a static image into a video with appealing 3D visual effects. Existing approaches typically first conduct monocular depth estimation, then render the input frame to subsequent frames with various viewpoints, and…

Computer Vision and Pattern Recognition · Computer Science 2023-02-22 Xiaodong Wang , Chenfei Wu , Shengming Yin , Minheng Ni , Jianfeng Wang , Linjie Li , Zhengyuan Yang , Fan Yang , Lijuan Wang , Zicheng Liu , Yuejian Fang , Nan Duan

3D scene reconstruction is essential for applications in virtual reality, robotics, and autonomous driving, enabling machines to understand and interact with complex environments. Traditional 3D Gaussian Splatting techniques rely on images…

Graphics · Computer Science 2025-03-04 Changlin Song , Jiaqi Wang , Liyun Zhu , He Weng

Humans exhibit an innate capacity to rapidly perceive and segment objects from video observations, and even mentally assemble them into structured 3D scenes. Replicating such capability, termed compositional 3D reconstruction, is pivotal…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Mingyu Dong , Chong Xia , Mingyuan Jia , Weichen Lyu , Long Xu , Zheng Zhu , Yueqi Duan

In this paper, we report on a parallel freeviewpoint video synthesis algorithm that can efficiently reconstruct a high-quality 3D scene representation of sports scenes. The proposed method focuses on a scene that is captured by multiple…

Computer Vision and Pattern Recognition · Computer Science 2019-08-01 Jun Chen , Ryosuke Watanabe , Keisuke Nonaka , Tomoaki Konno , Hiroshi Sankoh , Sei Naito

We introduce the first data-driven multi-view 3D point tracker, designed to track arbitrary points in dynamic scenes using multiple camera views. Unlike existing monocular trackers, which struggle with depth ambiguities and occlusion, or…

Computer Vision and Pattern Recognition · Computer Science 2025-08-29 Frano Rajič , Haofei Xu , Marko Mihajlovic , Siyuan Li , Irem Demir , Emircan Gündoğdu , Lei Ke , Sergey Prokudin , Marc Pollefeys , Siyu Tang

We explore the task of embodied view synthesis from monocular videos of deformable scenes. Given a minute-long RGBD video of people interacting with their pets, we render the scene from novel camera trajectories derived from the in-scene…

Computer Vision and Pattern Recognition · Computer Science 2023-10-03 Chonghyuk Song , Gengshan Yang , Kangle Deng , Jun-Yan Zhu , Deva Ramanan

Recent advances in illumination control extend image-based methods to video, yet still facing a trade-off between lighting fidelity and temporal consistency. Moving beyond relighting, a key step toward generative modeling of real-world…

Computer Vision and Pattern Recognition · Computer Science 2025-12-16 Tianqi Liu , Zhaoxi Chen , Zihao Huang , Shaocong Xu , Saining Zhang , Chongjie Ye , Bohan Li , Zhiguo Cao , Wei Li , Hao Zhao , Ziwei Liu

Estimating 3D scene flow from a sequence of monocular images has been gaining increased attention due to the simple, economical capture setup. Owing to the severe ill-posedness of the problem, the accuracy of current methods has been…

Computer Vision and Pattern Recognition · Computer Science 2021-05-06 Junhwa Hur , Stefan Roth

Sparse-view 3D reconstruction is essential for modeling scenes from casual captures, but remain challenging for non-generative reconstruction. Existing diffusion-based approaches mitigates this issues by synthesizing novel views, but they…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Yutian Chen , Shi Guo , Renbiao Jin , Tianshuo Yang , Xin Cai , Yawen Luo , Mingxin Yang , Mulin Yu , Linning Xu , Tianfan Xue

Camera control has been extensively studied in conditioned video generation; however, performing precisely altering the camera trajectories while faithfully preserving the video content remains a challenging task. The mainstream approach to…

Computer Vision and Pattern Recognition · Computer Science 2026-02-02 Dong-Yu Chen , Yixin Guo , Shuojin Yang , Tai-Jiang Mu , Shi-Min Hu

We present a solution for the goal of extracting a video from a single motion blurred image to sequentially reconstruct the clear views of a scene as beheld by the camera during the time of exposure. We first learn motion representation…

Computer Vision and Pattern Recognition · Computer Science 2022-01-31 Kuldeep Purohit , Anshul Shah , A. N. Rajagopalan

We present a fully automatic approach to real-time 3D face reconstruction from monocular in-the-wild videos. With the use of a cascaded-regressor based face tracking and a 3D Morphable Face Model shape fitting, we obtain a semi-dense 3D…

Computer Vision and Pattern Recognition · Computer Science 2017-08-28 Patrik Huber , Philipp Kopp , Matthias Rätsch , William Christmas , Josef Kittler

Panoramic video generation aims to synthesize 360-degree immersive videos, holding significant importance in the fields of VR, world models, and spatial intelligence. Existing works fail to synthesize high-quality panoramic videos due to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Zixun Fang , Kai Zhu , Zhiheng Liu , Yu Liu , Wei Zhai , Yang Cao , Zheng-Jun Zha
‹ Prev 1 4 5 6 7 8 10 Next ›