English
Related papers

Related papers: SpatialMe: Stereo Video Conversion Using Depth-War…

200 papers

Stereo video inpainting, which aims to fill the occluded regions of warped videos with visually coherent content while maintaining temporal consistency, remains a challenging open problem. The regions to be filled are scattered along object…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Yuan Huang , Sijie Zhao , Jing Cheng , Hao Xu , Shaohui Jiao

Stereo video synthesis from a monocular input is a demanding task in the fields of spatial computing and virtual reality. The main challenges of this task lie on the insufficiency of high-quality paired stereo videos for training and the…

Computer Vision and Pattern Recognition · Computer Science 2025-04-29 Zhen Lv , Yangqi Long , Congzhentao Huang , Cao Li , Chengfei Lv , Hao Ren , Dian Zheng

We tackle the problem of monocular-to-stereo video conversion and propose a novel architecture for inpainting and refinement of the warped right view obtained by depth-based reprojection of the input left view. We extend the Stable Video…

Computer Vision and Pattern Recognition · Computer Science 2025-11-26 Nina Shvetsova , Goutam Bhat , Prune Truong , Hilde Kuehne , Federico Tombari

This paper presents a novel framework for converting 2D videos to immersive stereoscopic 3D, addressing the growing demand for 3D content in immersive experience. Leveraging foundation models as priors, our approach overcomes the…

Computer Vision and Pattern Recognition · Computer Science 2024-09-12 Sijie Zhao , Wenbo Hu , Xiaodong Cun , Yong Zhang , Xiaoyu Li , Zhe Kong , Xiangjun Gao , Muyao Niu , Ying Shan

Stereo video generation has been gaining increasing attention with recent advancements in video diffusion models. However, most existing methods focus on generating 3D stereoscopic videos from monocular 2D videos. These approaches typically…

Computer Vision and Pattern Recognition · Computer Science 2025-06-09 Xingchang Huang , Ashish Kumar Singh , Florian Dubost , Cristina Nader Vasconcelos , Sakar Khattar , Liang Shi , Christian Theobalt , Cengiz Oztireli , Gurprit Singh

We introduce \textit{ImmersePro}, an innovative framework specifically designed to transform single-view videos into stereo videos. This framework utilizes a novel dual-branch architecture comprising a disparity branch and a context branch…

Computer Vision and Pattern Recognition · Computer Science 2024-10-02 Jian Shi , Zhenyu Li , Peter Wonka

While video generation models excel at producing high-quality monocular videos, generating 3D stereoscopic and spatial videos for immersive applications remains an underexplored challenge. We present a pose-free and training-free method…

Computer Vision and Pattern Recognition · Computer Science 2025-08-12 Peng Dai , Feitong Tan , Qiangeng Xu , Yihua Huang , David Futschik , Ruofei Du , Sean Fanello , Yinda Zhang , Xiaojuan Qi

We consider the problem of reconstructing a dynamic scene observed from a stereo camera. Most existing methods for depth from stereo treat different stereo frames independently, leading to temporally inconsistent depth predictions. Temporal…

Computer Vision and Pattern Recognition · Computer Science 2023-05-04 Nikita Karaev , Ignacio Rocco , Benjamin Graham , Natalia Neverova , Andrea Vedaldi , Christian Rupprecht

We present StereoWorld, a camera-conditioned stereo world model that jointly learns appearance and binocular geometry for end-to-end stereo video generation.Unlike monocular RGB or RGBD approaches, StereoWorld operates exclusively within…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Yang-Tian Sun , Zehuan Huang , Yifan Niu , Lin Ma , Yan-Pei Cao , Yuewen Ma , Xiaojuan Qi

The growing adoption of XR devices has fueled strong demand for high-quality stereo video, yet its production remains costly and artifact-prone. To address this challenge, we present StereoWorld, an end-to-end framework that repurposes a…

Computer Vision and Pattern Recognition · Computer Science 2025-12-12 Ke Xing , Xiaojie Jin , Longfei Li , Yuyang Yin , Hanwen Liang , Guixun Luo , Chen Fang , Jue Wang , Konstantinos N. Plataniotis , Yao Zhao , Yunchao Wei

The view synthesis problem--generating novel views of a scene from known imagery--has garnered recent attention due in part to compelling applications in virtual and augmented reality. In this paper, we explore an intriguing scenario for…

Computer Vision and Pattern Recognition · Computer Science 2018-05-25 Tinghui Zhou , Richard Tucker , John Flynn , Graham Fyffe , Noah Snavely

The rapid growth of stereoscopic displays, including VR headsets and 3D cinemas, has led to increasing demand for high-quality stereo video content. However, producing 3D videos remains costly and complex, while automatic…

Computer Vision and Pattern Recognition · Computer Science 2025-12-19 Guibao Shen , Yihua Du , Wenhang Ge , Jing He , Chirui Chang , Donghao Zhou , Zhen Yang , Luozhou Wang , Xin Tao , Ying-Cong Chen

Whether to attract viewer attention to a particular object, give the impression of depth or simply reproduce human-like scene perception, shallow depth of field images are used extensively by professional and amateur photographers alike. To…

Computer Vision and Pattern Recognition · Computer Science 2019-10-01 Benjamin Busam , Matthieu Hog , Steven McDonagh , Gregory Slabaugh

Reconstructing spatially and temporally coherent videos from time-varying measurements is a fundamental challenge in many scientific domains. A major difficulty arises from the sparsity of measurements, which hinders accurate recovery of…

Computer Vision and Pattern Recognition · Computer Science 2025-06-11 Bingliang Zhang , Zihui Wu , Berthy T. Feng , Yang Song , Yisong Yue , Katherine L. Bouman

Generating high-quality stereo videos requires consistent depth perception and temporal coherence across frames. Despite advances in image and video synthesis using diffusion models, producing high-quality stereo videos remains a…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Jian Shi , Qian Wang , Zhenyu Li , Wenqing Cui , Ramzi Idoughi , Peter Wonka

This paper introduces Stereo Any Video, a powerful framework for video stereo matching. It can estimate spatially accurate and temporally consistent disparities without relying on auxiliary information such as camera poses or optical flow.…

Computer Vision and Pattern Recognition · Computer Science 2025-07-22 Junpeng Jing , Weixun Luo , Ye Mao , Krystian Mikolajczyk

Video stereo matching is the task of estimating consistent disparity maps from rectified stereo videos. There is considerable scope for improvement in both datasets and methods within this area. Recent learning-based methods often focus on…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Junpeng Jing , Ye Mao , Anlan Qiu , Krystian Mikolajczyk

Video inpainting fills in corrupted video content with plausible replacements. While recent advances in endoscopic video inpainting have shown potential for enhancing the quality of endoscopic videos, they mainly repair 2D visual…

Image and Video Processing · Electrical Eng. & Systems 2024-07-04 Francis Xiatian Zhang , Shuang Chen , Xianghua Xie , Hubert P. H. Shum

The growing demand for immersive 3D content calls for automated monocular-to-stereo video conversion. We present Elastic3D, a controllable, direct end-to-end method for upgrading a conventional video to a binocular one. Our approach, based…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Nando Metzger , Prune Truong , Goutam Bhat , Konrad Schindler , Federico Tombari

Three key challenges hinder the development of current deepfake video detection: (1) Temporal features can be complex and diverse: how can we identify general temporal artifacts to enhance model generalization? (2) Spatiotemporal models…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Zhiyuan Yan , Yandan Zhao , Shen Chen , Mingyi Guo , Xinghe Fu , Taiping Yao , Shouhong Ding , Li Yuan
‹ Prev 1 2 3 10 Next ›