中文
相关论文

相关论文: WALT3D: Generating Realistic Training Data from Ti…

200 篇论文

While data has certainly taken the center stage in computer vision in recent years, it can still be difficult to obtain in certain scenarios. In particular, acquiring ground truth 3D shapes of objects pictured in 2D images remains a…

计算机视觉与模式识别 · 计算机科学 2016-08-02 Joao Carreira , Sara Vicente , Lourdes Agapito , Jorge Batista

Object geometry is key information for robot manipulation. Yet, object reconstruction is a challenging task because cameras only capture partial observations of objects, especially when occlusion occurs. In this paper, we leverage two extra…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Minghan Zhu , Zhiyi Wang , Qihang Sun , Maani Ghaffari , Michael Posa

High-Fidelity 3D scene reconstruction plays a crucial role in autonomous driving by enabling novel data generation from existing datasets. This allows simulating safety-critical scenarios and augmenting training datasets without incurring…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Pou-Chun Kung , Skanda Harisha , Ram Vasudevan , Aline Eid , Katherine A. Skinner

In this work, we present a method for a probabilistic fusion of external depth and onboard proximity data to form a volumetric 3-D map of a robot's environment. We extend the Octomap framework to update a representation of the area around…

机器人学 · 计算机科学 2021-10-25 Matthew Strong , Caleb Escobedo , Alessandro Roncone

Existing shape estimation methods for deformable object manipulation suffer from the drawbacks of being off-line, model dependent, noise-sensitive or occlusion-sensitive, and thus are not appropriate for manipulation tasks requiring high…

机器人学 · 计算机科学 2018-09-27 Tao Han , Xuan Zhao , Peigen Sun , Jia Pan

Occlusions are a common occurrence in unconstrained face images. Single image 3D reconstruction from such face images often suffers from corruption due to the presence of occlusions. Furthermore, while a plurality of 3D reconstructions is…

计算机视觉与模式识别 · 计算机科学 2022-04-04 Rahul Dey , Vishnu Naresh Boddeti

Rendering dynamic 3D human from monocular videos is crucial for various applications such as virtual reality and digital entertainment. Most methods assume the people is in an unobstructed scene, while various objects may cause the…

计算机视觉与模式识别 · 计算机科学 2025-02-21 Jingrui Ye , Zongkai Zhang , Yujiao Jiang , Qingmin Liao , Wenming Yang , Zongqing Lu

Deep networks for visual recognition are known to leverage "easy to recognise" portions of objects such as faces and distinctive texture patterns. The lack of a holistic understanding of objects may increase fragility and overfitting. In…

计算机视觉与模式识别 · 计算机科学 2019-10-28 Ruth Fong , Andrea Vedaldi

Generation of 3D data by deep neural network has been attracting increasing attention in the research community. The majority of extant works resort to regular representations such as volumetric grids or collection of images; however, these…

计算机视觉与模式识别 · 计算机科学 2016-12-08 Haoqiang Fan , Hao Su , Leonidas Guibas

Reconstructing dynamic driving scenes is essential for developing autonomous systems through sensor-realistic simulation. Although recent methods achieve high-fidelity reconstructions, they either rely on costly human annotations for object…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Carl Lindström , Mahan Rafidashti , Maryam Fatemi , Lars Hammarstrand , Martin R. Oswald , Lennart Svensson

The visualization of temporal data on urban buildings, such as shadows, noise, and solar potential, plays a critical role in the analysis of dynamic urban phenomena. However, in dense and geographically constrained 3D urban environments,…

人机交互 · 计算机科学 2026-02-04 Roberta Mota , Julio D. Silva , Fabio Miranda , Usman Alim , Ehud Sharlin , Nivan Ferreira

Data augmentation plays a crucial role in deep learning, enhancing the generalization and robustness of learning-based models. Standard approaches involve simple transformations like rotations and flips for generating extra data. However,…

计算机视觉与模式识别 · 计算机科学 2024-08-27 Shichao Dong , Ze Yang , Guosheng Lin

The interest in 3D dynamical tracking is growing in fields such as robotics, biology and fluid dynamics. Recently, a major source of progress in 3D tracking has been the study of collective behaviour in biological systems, where the…

计算机视觉与模式识别 · 计算机科学 2015-11-05 Andrea Cavagna , Chiara Creato , Lorenzo Del Castello , Stefania Melillo , Leonardo Parisi , Massimiliano Viale

Recently, deep learning-based 3D face reconstruction methods have demonstrated promising advancements in terms of quality and efficiency. Nevertheless, these techniques face challenges in effectively handling occluded scenes and fail to…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Dapeng Zhao

A 3D scene consists of a set of objects, each with a shape and a layout giving their position in space. Understanding 3D scenes from 2D images is an important goal, with applications in robotics and graphics. While there have been recent…

计算机视觉与模式识别 · 计算机科学 2022-06-15 Georgia Gkioxari , Nikhila Ravi , Justin Johnson

Reconstructing 3D objects from a single image remains challenging, especially under real-world occlusions. While recent diffusion-based view synthesis models can generate consistent novel views from a single RGB image, they typically assume…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Yansong Qu , Shaohui Dai , Xinyang Li , Yuze Wang , You Shen , Liujuan Cao , Rongrong Ji

In autonomous driving, data augmentation is commonly used for improving 3D object detection. The most basic methods include insertion of copied objects and rotation and scaling of the entire training frame. Numerous variants have been…

计算机视觉与模式识别 · 计算机科学 2023-09-01 Jungwook Shin , Jaeill Kim , Kyungeun Lee , Hyunghun Cho , Wonjong Rhee

The capability to accurately estimate 3D human poses is crucial for diverse fields such as action recognition, gait recognition, and virtual/augmented reality. However, a persistent and significant challenge within this field is the…

Discovering 3D arrangements of objects from single indoor images is important given its many applications including interior design, content creation, etc. Although heavily researched in the recent years, existing approaches break down…

计算机视觉与模式识别 · 计算机科学 2017-12-05 Moos Hueting , Pradyumna Reddy , Vladimir Kim , Ersin Yumer , Nathan Carr , Niloy Mitra

Previous surface reconstruction methods either suffer from low geometric accuracy or lengthy training times when dealing with real-world complex dynamic scenes involving multi-person activities, and human-object interactions. To tackle the…

计算机视觉与模式识别 · 计算机科学 2024-09-30 Shuo Wang , Binbin Huang , Ruoyu Wang , Shenghua Gao