中文
相关论文

相关论文: OASIS: A Large-Scale Dataset for Single Image 3D i…

200 篇论文

Single-view depth prediction is a fundamental problem in computer vision. Recently, deep learning methods have led to significant progress, but such methods are limited by the available training data. Current datasets based on 3D sensors…

计算机视觉与模式识别 · 计算机科学 2018-11-29 Zhengqi Li , Noah Snavely

We study 3D shape modeling from a single image and make contributions to it in three aspects. First, we present Pix3D, a large-scale benchmark of diverse image-shape pairs with pixel-level 2D-3D alignment. Pix3D has wide applications in…

计算机视觉与模式识别 · 计算机科学 2018-04-13 Xingyuan Sun , Jiajun Wu , Xiuming Zhang , Zhoutong Zhang , Chengkai Zhang , Tianfan Xue , Joshua B. Tenenbaum , William T. Freeman

The Simple Image Access protocol (SIA) provides capabilities for the discovery, description, access, and retrieval of multi-dimensional image datasets, including 2-D images as well as datacubes of three or more dimensions. SIA data…

天体物理仪器与方法 · 物理学 2016-01-05 Patrick Dowler , Doug Tody , François Bonnarel

Single image 3D reconstruction is an important but challenging task that requires extensive knowledge of our natural world. Many existing methods solve this problem by optimizing a neural radiance field under the guidance of 2D diffusion…

计算机视觉与模式识别 · 计算机科学 2023-06-30 Minghua Liu , Chao Xu , Haian Jin , Linghao Chen , Mukund Varma T , Zexiang Xu , Hao Su

Recognizing scenes and objects in 3D from a single image is a longstanding goal of computer vision with applications in robotics and AR/VR. For 2D recognition, large datasets and scalable solutions have led to unprecedented advances. In 3D,…

计算机视觉与模式识别 · 计算机科学 2023-03-27 Garrick Brazil , Abhinav Kumar , Julian Straub , Nikhila Ravi , Justin Johnson , Georgia Gkioxari

Recent advances in leveraging large-scale Internet photo collections for 3D reconstruction have enabled immersive virtual exploration of landmarks and historic sites worldwide. However, little attention has been given to the immersive…

图形学 · 计算机科学 2025-08-06 Yuze Wang , Yue Qi

We present SAM 3D, a generative model for visually grounded 3D object reconstruction, predicting geometry, texture, and layout from a single image. SAM 3D excels in natural images, where occlusion and scene clutter are common and visual…

One major challenge for monocular 3D human pose estimation in-the-wild is the acquisition of training data that contains unconstrained images annotated with accurate 3D poses. In this paper, we address this challenge by proposing a…

计算机视觉与模式识别 · 计算机科学 2020-03-18 Umar Iqbal , Pavlo Molchanov , Jan Kautz

Single image 3D photography enables viewers to view a still image from novel viewpoints. Recent approaches combine monocular depth networks with inpainting networks to achieve compelling results. A drawback of these techniques is the use of…

The growing prevalence of intelligent manufacturing and autonomous vehicles has intensified the demand for three-dimensional (3D) reconstruction under complex reflection and transmission conditions. Traditional structured light techniques…

Monocular 3D object detection (M3OD) is intrinsically ill-posed, hence training a high-performance deep learning based M3OD model requires a humongous amount of labeled data with complicated visual variation from diverse scenes, variety of…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Zhaonian Kuang , Rui Ding , Meng Yang , Xinhu Zheng , Gang Hua

We introduce a new large-scale dataset for the advancement of object detection techniques and overhead object detection research. This satellite imagery dataset enables research progress pertaining to four key computer vision frontiers. We…

计算机视觉与模式识别 · 计算机科学 2018-02-23 Darius Lam , Richard Kuzma , Kevin McGee , Samuel Dooley , Michael Laielli , Matthew Klaric , Yaroslav Bulatov , Brendan McCord

Current perception models in autonomous driving have become notorious for greatly relying on a mass of annotated data to cover unseen cases and address the long-tail problem. On the other hand, learning from unlabeled large-scale collected…

计算机视觉与模式识别 · 计算机科学 2021-10-26 Jiageng Mao , Minzhe Niu , Chenhan Jiang , Hanxue Liang , Jingheng Chen , Xiaodan Liang , Yamin Li , Chaoqiang Ye , Wei Zhang , Zhenguo Li , Jie Yu , Hang Xu , Chunjing Xu

Reconstructing and rendering 3D objects from highly sparse views is of critical importance for promoting applications of 3D vision techniques and improving user experience. However, images from sparse views only contain very limited 3D…

计算机视觉与模式识别 · 计算机科学 2024-11-14 Chen Yang , Sikuang Li , Jiemin Fang , Ruofan Liang , Lingxi Xie , Xiaopeng Zhang , Wei Shen , Qi Tian

We propose a 3D novel sparse-view synthesis framework for unconstrained real-world scenarios that contain distractors. Unlike existing methods that primarily perform novel-view synthesis from a sparse set of constrained images without…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Wongi Park , Jordan A. James , Myeongseok Nam , Minjae Lee , Soomok Lee , Sang-Hyun Lee , William J. Beksi

Object reconstruction from a single image -- in the wild -- is a problem where we can make progress and get meaningful results today. This is the main message of this paper, which introduces an automated pipeline with pixels as inputs and…

计算机视觉与模式识别 · 计算机科学 2015-05-08 Abhishek Kar , Shubham Tulsiani , João Carreira , Jitendra Malik

The idea of 3D reconstruction as scene understanding is foundational in computer vision. Reconstructing 3D scenes from 2D visual observations requires strong priors to disambiguate structure. Much work has been focused on the…

计算机视觉与模式识别 · 计算机科学 2024-12-02 Peter Kulits , Michael J. Black , Silvia Zuffi

In the past decade, object detection has achieved significant progress in natural images but not in aerial images, due to the massive variations in the scale and orientation of objects caused by the bird's-eye view of aerial images. More…

计算机视觉与模式识别 · 计算机科学 2021-12-07 Jian Ding , Nan Xue , Gui-Song Xia , Xiang Bai , Wen Yang , Micheal Ying Yang , Serge Belongie , Jiebo Luo , Mihai Datcu , Marcello Pelillo , Liangpei Zhang

Implicit neural representation methods have shown impressive advancements in learning 3D scenes from unstructured in-the-wild photo collections but are still limited by the large computational cost of volumetric rendering. More recently, 3D…

计算机视觉与模式识别 · 计算机科学 2024-04-08 Hiba Dahmani , Moussab Bennehar , Nathan Piasco , Luis Roldao , Dzmitry Tsishkou