English
Related papers

Related papers: Towards Learning a Generalizable 3D Scene Represen…

200 papers

Accurate 3D scene understanding is essential for embodied intelligence, with occupancy prediction emerging as a key task for reasoning about both objects and free space. Existing approaches largely rely on depth priors (e.g., DepthAnything)…

Computer Vision and Pattern Recognition · Computer Science 2026-02-26 Changqing Zhou , Yueru Luo , Changhao Chen

One essential step to realize modern driver assistance technology is the accurate knowledge about the location of static objects in the environment. In this work, we use artificial neural networks to predict the occupation state of a whole…

Robotics · Computer Science 2019-04-01 Daniel Bauer , Lars Kuhnert , Lutz Eckstein

For humans, visual understanding is inherently generative: given a 3D shape, we can postulate how it would look in the world; given a 2D image, we can infer the 3D structure that likely gave rise to it. We can thus translate between the 2D…

Computer Vision and Pattern Recognition · Computer Science 2020-11-17 Tristan Aumentado-Armstrong , Alex Levinshtein , Stavros Tsogkas , Konstantinos G. Derpanis , Allan D. Jepson

Unsupervised learning with generative models has the potential of discovering rich representations of 3D scenes. While geometric deep learning has explored 3D-structure-aware representations of scene geometry, these models typically require…

Computer Vision and Pattern Recognition · Computer Science 2020-01-30 Vincent Sitzmann , Michael Zollhöfer , Gordon Wetzstein

We propose a method for estimating the 6DoF pose of a rigid object with an available 3D model from a single RGB image. Unlike classical correspondence-based methods which predict 3D object coordinates at pixels of the input image, the…

Computer Vision and Pattern Recognition · Computer Science 2022-08-02 Lin Huang , Tomas Hodan , Lingni Ma , Linguang Zhang , Luan Tran , Christopher Twigg , Po-Chen Wu , Junsong Yuan , Cem Keskin , Robert Wang

We present dynamic neural radiance fields for modeling the appearance and dynamics of a human face. Digitally modeling and reconstructing a talking human is a key building-block for a variety of applications. Especially, for telepresence…

Computer Vision and Pattern Recognition · Computer Science 2020-12-08 Guy Gafni , Justus Thies , Michael Zollhöfer , Matthias Nießner

We consider the challenging problem of outdoor lighting estimation for the goal of photorealistic virtual object insertion into photographs. Existing works on outdoor lighting estimation typically simplify the scene lighting into an…

Computer Vision and Pattern Recognition · Computer Science 2022-08-22 Zian Wang , Wenzheng Chen , David Acuna , Jan Kautz , Sanja Fidler

In this paper we provide an overview of a new framework for robot perception, real-world modelling, and navigation that uses a stochastic tesselated representation of spatial information called the Occupancy Grid. The Occupancy Grid is a…

Robotics · Computer Science 2013-04-05 A. Elfes

3D occupancy prediction aims to infer dense, voxel-wise scene semantics from sensor observations, where the 2D-to-3D view transformation serves as a crucial step in bridging image features and volumetric representations. Most previous…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Yuan Wu , Zhiqiang Yan , Jiawei Lian , Zhengxue Wang , Jian Yang

Utilizing multi-view inputs to synthesize novel-view images, Neural Radiance Fields (NeRF) have emerged as a popular research topic in 3D vision. In this work, we introduce a Generalizable Semantic Neural Radiance Field (GSNeRF), which…

Computer Vision and Pattern Recognition · Computer Science 2024-03-07 Zi-Ting Chou , Sheng-Yu Huang , I-Jieh Liu , Yu-Chiang Frank Wang

Unsupervised learning of 3D-aware generative adversarial networks has lately made much progress. Some recent work demonstrates promising results of learning human generative models using neural articulated radiance fields, yet their…

Computer Vision and Pattern Recognition · Computer Science 2023-09-12 Xinya Chen , Jiaxin Huang , Yanrui Bin , Lu Yu , Yiyi Liao

Global localization of a mobile robot using planar surface segments extracted from depth images is considered. The robot's environment is represented by a topological map consisting of local models, each representing a particular location…

Computer Vision and Pattern Recognition · Computer Science 2013-10-02 Robert Cupec , Emmanuel Karlo Nyarko , Damir Filko , Andrej Kitanov , Ivan Petrović

State-of-the-art navigation methods leverage a spatial memory to generalize to new environments, but their occupancy maps are limited to capturing the geometric structures directly observed by the agent. We propose occupancy anticipation,…

Computer Vision and Pattern Recognition · Computer Science 2020-08-26 Santhosh K. Ramakrishnan , Ziad Al-Halah , Kristen Grauman

We present a method for composing photorealistic scenes from captured images of objects. Our work builds upon neural radiance fields (NeRFs), which implicitly model the volumetric density and directionally-emitted radiance of a scene. While…

Computer Vision and Pattern Recognition · Computer Science 2020-12-16 Michelle Guo , Alireza Fathi , Jiajun Wu , Thomas Funkhouser

We address efficient and structure-aware 3D scene representation from images. Nerflets are our key contribution -- a set of local neural radiance fields that together represent a scene. Each nerflet maintains its own spatial position,…

Computer Vision and Pattern Recognition · Computer Science 2023-03-13 Xiaoshuai Zhang , Abhijit Kundu , Thomas Funkhouser , Leonidas Guibas , Hao Su , Kyle Genova

Neural radiance fields, which represent a 3D scene as a color field and a density field, have demonstrated great progress in novel view synthesis yet are unfavorable for editing due to the implicitness. This work studies the task of…

Computer Vision and Pattern Recognition · Computer Science 2025-02-14 Ka Leong Cheng , Qiuyu Wang , Zifan Shi , Kecheng Zheng , Yinghao Xu , Hao Ouyang , Qifeng Chen , Yujun Shen

Learning robot control policies from human videos is a promising direction for scaling up robot learning. However, how to extract action knowledge (or action representations) from videos for policy learning remains a key challenge. Existing…

Robotics · Computer Science 2025-06-05 Zhao-Heng Yin , Sherry Yang , Pieter Abbeel

Planetary exploration increasingly relies on autonomous robotic systems capable of perceiving, interpreting, and reconstructing their surroundings in the absence of global positioning or real-time communication with Earth. Rovers operating…

3D occupancy prediction provides a comprehensive description of the surrounding scenes and has become an essential task for 3D perception. Most existing methods focus on offline perception from one or a few views and cannot be applied to…

Computer Vision and Pattern Recognition · Computer Science 2025-08-26 Yuqi Wu , Wenzhao Zheng , Sicheng Zuo , Yuanhui Huang , Jie Zhou , Jiwen Lu

3D scene segmentation based on neural implicit representation has emerged recently with the advantage of training only on 2D supervision. However, existing approaches still requires expensive per-scene optimization that prohibits…

Computer Vision and Pattern Recognition · Computer Science 2023-10-27 Hanlin Chen , Chen Li , Mengqi Guo , Zhiwen Yan , Gim Hee Lee
‹ Prev 1 8 9 10 Next ›