English
Related papers

Related papers: OASIS: A Large-Scale Dataset for Single Image 3D i…

200 papers

Simulation systems have become an essential component in the development and validation of autonomous driving technologies. The prevailing state-of-the-art approach for simulation is to use game engines or high-fidelity computer graphics…

Computer Vision and Pattern Recognition · Computer Science 2020-10-30 Wei Li , Chengwei Pan , Rong Zhang , Jiaping Ren , Yuexin Ma , Jin Fang , Feilong Yan , Qichuan Geng , Xinyu Huang , Huajun Gong , Weiwei Xu , Guoping Wang , Dinesh Manocha , Ruigang Yang

With deep learning and computer vision technology development, autonomous driving provides new solutions to improve traffic safety and efficiency. The importance of building high-quality datasets is self-evident, especially with the rise of…

Computer Vision and Pattern Recognition · Computer Science 2024-02-07 Guohang Yan , Jiahao Pi , Jianfei Guo , Zhaotong Luo , Min Dou , Nianchen Deng , Qiusheng Huang , Daocheng Fu , Licheng Wen , Pinlong Cai , Xing Gao , Xinyu Cai , Bo Zhang , Xuemeng Yang , Yeqi Bai , Hongbin Zhou , Botian Shi

A long-standing goal of 3D human reconstruction is to create lifelike and fully detailed 3D humans from single-view images. The main challenge lies in inferring unknown body shapes, appearances, and clothing details in areas not visible in…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Hsuan-I Ho , Jie Song , Otmar Hilliges

Dense 3D reconstruction and tracking of dynamic scenes from monocular video remains an important open challenge in computer vision. Progress in this area has been constrained by the scarcity of high-quality datasets with dense, complete,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-07 Zeren Jiang , Yushi Lan , Yihang Luo , Yufan Deng , Zihang Lai , Edgar Sucar , Christian Rupprecht , Iro Laina , Diane Larlus , Chuanxia Zheng , Andrea Vedaldi

Robust perception is critical for autonomous driving, especially under adverse weather and lighting conditions that commonly occur in real-world environments. In this paper, we introduce the Stereo Image Dataset (SID), a large-scale…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Zaid A. El-Shair , Abdalmalek Abu-raddaha , Aaron Cofield , Hisham Alawneh , Mohamed Aladem , Yazan Hamzeh , Samir A. Rawashdeh

We present a novel data set made up of omnidirectional video of multiple objects whose centroid positions are annotated automatically. Omnidirectional vision is an active field of research focused on the use of spherical imagery in video…

Computer Vision and Pattern Recognition · Computer Science 2017-09-13 Victor Stamatescu , Peter Barsznica , Manjung Kim , Kin K. Liu , Mark McKenzie , Will Meakin , Gwilyn Saunders , Sebastien C. Wong , Russell S. A. Brinkworth

3D reconstruction is a fundamental task in robotics that gained attention due to its major impact in a wide variety of practical settings, including agriculture, underwater, and urban environments. This task can be carried out via view…

We introduce the S-EO dataset: a large-scale, high-resolution dataset, designed to advance geometry-aware shadow detection. Collected from diverse public-domain sources, including challenge datasets and government providers such as USGS,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Elías Masquil , Roger Marí , Thibaud Ehret , Enric Meinhardt-Llopis , Pablo Musé , Gabriele Facciolo

Humans can easily understand a single image as depicting multiple potential objects permitting interaction. We use this skill to plan our interactions with the world and accelerate understanding new objects without engaging in interaction.…

Computer Vision and Pattern Recognition · Computer Science 2023-08-09 Shengyi Qian , David F. Fouhey

We propose VASA-3D, an audio-driven, single-shot 3D head avatar generator. This research tackles two major challenges: capturing the subtle expression details present in real human faces, and reconstructing an intricate 3D head avatar from…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Sicheng Xu , Guojun Chen , Jiaolong Yang , Yizhong Zhang , Yu Deng , Steve Lin , Baining Guo

Imaging through thick scattering media presents significant challenges, particularly for three-dimensional (3D) applications. This manuscript demonstrates a novel scheme for single-image-enabled 3D imaging through such media, treating the…

Optics · Physics 2024-10-10 Long Pan , Yunan Wang , Yijie Lou , Xiaohua Feng

Production of photorealistic, navigable 3D site models requires a large volume of carefully collected images that are often unavailable to first responders for disaster relief or law enforcement. Real-world challenges include limited…

Computer Vision and Pattern Recognition · Computer Science 2025-05-05 Neil Joshi , Joshua Carney , Nathanael Kuo , Homer Li , Cheng Peng , Myron Brown

Open-world perception aims to develop a model adaptable to novel domains and various sensor configurations and can understand uncommon objects and corner cases. However, current research lacks sufficiently comprehensive open-world 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Zhongyu Xia , Jishuo Li , Zhiwei Lin , Xinhao Wang , Yongtao Wang , Ming-Hsuan Yang

The lack of spatial dimensional information remains a challenge in normal estimation from a single image. Recent diffusion-based methods have demonstrated significant potential in 2D-to-3D implicit mapping, they rely on data-driven…

Computer Vision and Pattern Recognition · Computer Science 2025-08-11 Yanxing Liang , Yinghui Wang , Jinlong Yang , Wei Li

Existing Earth Vision datasets are either suitable for semantic segmentation or object detection. In this work, we introduce the first benchmark dataset for instance segmentation in aerial imagery that combines instance-level object…

Computer Vision and Pattern Recognition · Computer Science 2019-08-29 Syed Waqas Zamir , Aditya Arora , Akshita Gupta , Salman Khan , Guolei Sun , Fahad Shahbaz Khan , Fan Zhu , Ling Shao , Gui-Song Xia , Xiang Bai

The pursuit of autonomous driving has produced one of the richest sensor data collections in all of robotics. However, its scale and diversity remain largely untapped. Each dataset adopts different 2D and 3D modalities, such as cameras,…

As many different 3D volumes could produce the same 2D x-ray image, inverting this process is challenging. We show that recent deep learning-based convolutional neural networks can solve this task. As the main challenge in learning is the…

Graphics · Computer Science 2018-11-29 Philipp Henzler , Volker Rasche , Timo Ropinski , Tobias Ritschel

Demand for image editing has been increasing as users' desire for expression is also increasing. However, for most users, image editing tools are not easy to use since the tools require certain expertise in photo effects and have complex…

Computation and Language · Computer Science 2022-02-25 Hyounghun Kim , Doo Soon Kim , Seunghyun Yoon , Franck Dernoncourt , Trung Bui , Mohit Bansal

Despite the rapid progress in data-driven 3D vision, aerial geometric 3D vision remains a formidable challenge due to the severe scarcity of large-scale, high-fidelity training data. Existing benchmarks, predominantly biased toward…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Xiaoya Cheng , Rouwan Wu , Xinyi Liu , Zeyu Cui , Yan Liu , Na Zhao , Yu Liu , Maojun Zhang , Shen Yan

Machine learning based Single Image Intrinsic Decomposition (SIID) methods decompose a captured scene into its albedo and shading images by using the knowledge of a large set of known and realistic ground truth decompositions. Collecting…

Computer Vision and Pattern Recognition · Computer Science 2018-09-05 Louis Lettry , Kenneth Vanhoey , Luc van Gool