English
Related papers

Related papers: StereoGenBench: A Synthetic Multi-Camera Benchmark…

200 papers

Methods and datasets for human pose estimation focus predominantly on side- and front-view scenarios. We overcome the limitation by leveraging synthetic data and introduce RePoGen (RarE POses GENerator), an SMPL-based method for generating…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Miroslav Purkrabek , Jiri Matas

Generation of images containing multiple humans, performing complex actions, while preserving their facial identities, is a significant challenge. A major factor contributing to this is the lack of a dedicated benchmark. To address this, we…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Shubhankar Borse , Seokeon Choi , Sunghyun Park , Jeongho Kim , Shreya Kadambi , Risheek Garrepalli , Sungrack Yun , Munawar Hayat , Fatih Porikli

Incorporating camera intrinsics into video generation models offers a principled way to control not only scene dynamics but also the imaging process that governs visual appearance. Prior work has primarily focused on extrinsic control, such…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Debabrata Mandal , Zhihan Peng , Yujie Wang , Praneeth Chakravarthula

Closed-loop simulation and scalable pre-training for autonomous driving require synthesizing free-viewpoint driving scenes. However, existing datasets and generative pipelines rarely provide consistent off-trajectory observations, limiting…

Computer Vision and Pattern Recognition · Computer Science 2025-12-05 Shijie Chen , Peixi Peng

The rising popularity of immersive visual experiences has increased interest in stereoscopic 3D video generation. Despite significant advances in video synthesis, creating 3D videos remains challenging due to the relative scarcity of 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-05-02 Michal Geyer , Omer Tov , Linyi Jin , Richard Tucker , Inbar Mosseri , Tali Dekel , Noah Snavely

It is well known that the passive stereo system cannot adapt well to weak texture objects, e.g., white walls. However, these weak texture targets are very common in indoor environments. In this paper, we present a novel stereo system, which…

Computer Vision and Pattern Recognition · Computer Science 2022-03-22 Yuhua Xu , Xiaoli Yang , Yushan Yu , Wei Jia , Zhaobi Chu , Yulan Guo

For a machine learning model to generalize effectively to unseen data within a particular problem domain, it is well-understood that the data needs to be of sufficient size and representative of real-world scenarios. Nonetheless, real-world…

Computer Vision and Pattern Recognition · Computer Science 2023-08-08 Kidist Amde Mekonnen

Semantic segmentation is an essential step for many vision applications in order to understand a scene and the objects within. Recent progress in hyperspectral imaging technology enables the application in driving scenarios and the hope is…

Computer Vision and Pattern Recognition · Computer Science 2025-09-18 Nick Theisen , Robin Bartsch , Dietrich Paulus , Peer Neubert

Spatial computing experiences are constrained by the real-world surroundings of the user. In such experiences, augmenting virtual objects to existing scenes require a contextual approach, where geometrical conflicts are avoided, and…

Graphics · Computer Science 2020-10-01 Mohammad Keshavarzi , Aakash Parikh , Xiyu Zhai , Melody Mao , Luisa Caldas , Allen Y. Yang

This research paper introduces a synthetic hyperspectral dataset that combines high spectral and spatial resolution imaging to achieve a comprehensive, accurate, and detailed representation of observed scenes or objects. Obtaining such…

Computer Vision and Pattern Recognition · Computer Science 2023-09-04 Yajie Sun , Ali Zia , Jun Zhou

Photometric stereo typically demands intricate data acquisition setups involving multiple light sources to recover surface normals accurately. In this paper, we propose MERLiN, an attention-based hourglass network that integrates single…

Computer Vision and Pattern Recognition · Computer Science 2024-09-04 Ashish Tiwari , Satoshi Ikehata , Shanmuganathan Raman

Deep stereo matching has advanced significantly on benchmark datasets through fine-tuning but falls short of the zero-shot generalization seen in foundation models in other vision tasks. We introduce CogStereo, a novel framework that…

Computer Vision and Pattern Recognition · Computer Science 2025-10-28 Lihuang Fang , Xiao Hu , Yuchen Zou , Hong Zhang

We introduce MultiDiff, a novel approach for consistent novel view synthesis of scenes from a single RGB image. The task of synthesizing novel views from a single reference image is highly ill-posed by nature, as there exist multiple,…

Computer Vision and Pattern Recognition · Computer Science 2024-06-27 Norman Müller , Katja Schwarz , Barbara Roessle , Lorenzo Porzi , Samuel Rota Bulò , Matthias Nießner , Peter Kontschieder

Recent advances in robot imitation learning have yielded powerful visuomotor policies capable of manipulating a wide variety of objects directly from monocular visual inputs. However, monocular observations inherently lack reliable depth…

Robotics · Computer Science 2026-05-12 Evans Han , Yunfan Jiang , Yingke Wang , Haoyue Xiao , Huang Huang , Jianwen Xie , Jiajun Wu , Li Fei-Fei , Ruohan Zhang

Recent advancements in video diffusion models have shown exceptional abilities in simulating real-world dynamics and maintaining 3D consistency. This progress inspires us to investigate the potential of these models to ensure dynamic…

Computer Vision and Pattern Recognition · Computer Science 2024-12-11 Jianhong Bai , Menghan Xia , Xintao Wang , Ziyang Yuan , Xiao Fu , Zuozhu Liu , Haoji Hu , Pengfei Wan , Di Zhang

We design a multiscopic vision system that utilizes a low-cost monocular RGB camera to acquire accurate depth estimation. Unlike multi-view stereo with images captured at unconstrained camera poses, the proposed system controls the motion…

Computer Vision and Pattern Recognition · Computer Science 2021-08-21 Weihao Yuan , Rui Fan , Michael Yu Wang , Qifeng Chen

High-resolution (5MP+) stereo vision systems are essential for advancing robotic capabilities, enabling operation over longer ranges and generating significantly denser and accurate 3D point clouds. However, realizing the full potential of…

Geometry- and appearance-controlled full-body human image generation is an interesting but challenging task. Existing solutions are either unconditional or dependent on coarse conditions (e.g., pose, text), thus lacking explicit geometry…

Computer Vision and Pattern Recognition · Computer Science 2024-04-25 Linzi Qu , Jiaxiang Shang , Hui Ye , Xiaoguang Han , Hongbo Fu

Learning-based monocular depth estimation leverages geometric priors present in the training data to enable metric depth perception from a single image, a traditionally ill-posed problem. However, these priors are often specific to a…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Karlo Koledić , Luka Petrović , Ivan Petrović , Ivan Marković

Synthetic image datasets offer unmatched advantages for designing and evaluating deep neural networks: they make it possible to (i) render as many data samples as needed, (ii) precisely control each scene and yield granular ground truth…

Computer Vision and Pattern Recognition · Computer Science 2023-12-14 Florian Bordes , Shashank Shekhar , Mark Ibrahim , Diane Bouchacourt , Pascal Vincent , Ari S. Morcos