English
Related papers

Related papers: Next-Best View Policy for 3D Reconstruction

200 papers

Can VLMs predict how each camera move changes the view, and plan many such moves ahead? We call this capability view planning, requiring (1)understanding how a single action transforms the view, and (2)composing many such transformations…

Artificial Intelligence · Computer Science 2026-05-29 Kangrui Wang , Linjie Li , Zhengyuan Yang , Shiqi Chen , Zihan Wang , Li Fei-Fei , Jiajun Wu , Leonidas Guibas , Lijuan Wang , Manling Li

High dimensional data analysis for exploration and discovery includes three fundamental tasks: dimensionality reduction, clustering, and visualization. When the three associated tasks are done separately, as is often the case thus far,…

Machine Learning · Computer Science 2020-12-02 Stan Z. Li , Lirong Wu , Zelin Zang

Automatic reconstruction of 3D models from images using multi-view Structure-from-Motion methods has been one of the most fruitful outcomes of computer vision. These advances combined with the growing popularity of Micro Aerial Vehicles as…

Robotics · Computer Science 2016-11-15 Shreyansh Daftry , Christof Hoppe , Horst Bischof

We present an analysis of three possible strategies for exploiting the power of existing convolutional neural networks (ConvNets) in different scenarios from the ones they were trained: full training, fine tuning, and using ConvNets as…

Computer Vision and Pattern Recognition · Computer Science 2016-11-08 Keiller Nogueira , Otávio A. B. Penatti , Jefersson A. dos Santos

This paper introduces InverseMatrixVT3D, an efficient method for transforming multi-view image features into 3D feature volumes for 3D semantic occupancy prediction. Existing methods for constructing 3D volumes often rely on depth…

Computer Vision and Pattern Recognition · Computer Science 2024-04-30 Zhenxing Ming , Julie Stephany Berrio , Mao Shan , Stewart Worrall

Single-image piece-wise planar 3D reconstruction aims to simultaneously segment plane instances and recover 3D plane parameters from an image. Most recent approaches leverage convolutional neural networks (CNNs) and achieve promising…

Computer Vision and Pattern Recognition · Computer Science 2019-04-25 Zehao Yu , Jia Zheng , Dongze Lian , Zihan Zhou , Shenghua Gao

Object reconstruction is relevant for many autonomous robotic tasks that require interaction with the environment. A key challenge in such scenarios is planning view configurations to collect informative measurements for reconstructing an…

Robotics · Computer Science 2024-09-17 Sicong Pan , Liren Jin , Xuying Huang , Cyrill Stachniss , Marija Popović , Maren Bennewitz

We introduce a method to classify imagery using a convo- lutional neural network (CNN) on multi-view image pro- jections. The power of our method comes from using pro- jections of multiple images at multiple depth planes near the…

Computer Vision and Pattern Recognition · Computer Science 2017-12-27 Dror Aiger , Brett Allen , Aleksey Golovinskiy

3D building reconstruction from monocular remote sensing images is an important and challenging research problem that has received increasing attention in recent years, owing to its low cost of data acquisition and availability for…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Weijia Li , Haote Yang , Zhenghao Hu , Juepeng Zheng , Gui-Song Xia , Conghui He

Recent medical image reconstruction techniques focus on generating high-quality medical images suitable for clinical use at the lowest possible cost and with the fewest possible adverse effects on patients. Recent works have shown…

Image and Video Processing · Electrical Eng. & Systems 2025-01-07 Shijun Liang , Anish Lahiri , Saiprasad Ravishankar

We develop a framework for task-specific active next-best-view selection in 3D reconstruction from point clouds, by casting the problem in the language of Bayesian decision theory. Our framework works by (a) placing a prior distribution…

Graphics · Computer Science 2026-05-07 Jingsen Zhu , Silvia Sellán , Alexander Terenin

Autonomous visual navigation is an essential element in robot autonomy. Reinforcement learning (RL) offers a promising policy training paradigm. However existing RL methods suffer from high sample complexity, poor sim-to-real transfer, and…

Robotics · Computer Science 2025-07-31 Qianzhong Chen , Jiankai Sun , Naixiang Gao , JunEn Low , Timothy Chen , Mac Schwager

Active vision is inherently attention-driven: The agent actively selects views to attend in order to fast achieve the vision task while improving its internal representation of the scene being observed. Inspired by the recent success of…

Computer Vision and Pattern Recognition · Computer Science 2022-01-12 Min Liu , Yifei Shi , Lintao Zheng , Kai Xu , Hui Huang , Dinesh Manocha

Novel view synthesis (NVS) enables to generate new images of a scene or convert a set of 2D images into a comprehensive 3D model. In the context of Space Domain Awareness, since space is becoming increasingly congested, NVS can accurately…

Computer Vision and Pattern Recognition · Computer Science 2024-10-08 Nidhi Mathihalli , Audrey Wei , Giovanni Lavezzi , Peng Mun Siew , Victor Rodriguez-Fernandez , Hodei Urrutxua , Richard Linares

In this paper, we present a Convolutional Neural Network (CNN) regression approach for real-time 2-D/3-D registration. Different from optimization-based methods, which iteratively optimize the transformation parameters over a scalar-valued…

Computer Vision and Pattern Recognition · Computer Science 2016-04-26 Shun Miao , Z. Jane Wang , Rui Liao

We propose a novel approach to robot-operated active understanding of unknown indoor scenes, based on online RGBD reconstruction with semantic segmentation. In our method, the exploratory robot scanning is both driven by and targeting at…

Graphics · Computer Science 2022-01-14 Lintao Zheng , Chenyang Zhu , Jiazhao Zhang , Hang Zhao , Hui Huang , Matthias Niessner , Kai Xu

Recent studies on visual reinforcement learning (visual RL) have explored the use of 3D visual representations. However, none of these work has systematically compared the efficacy of 3D representations with 2D representations across…

Robotics · Computer Science 2023-06-13 Zhan Ling , Yunchao Yao , Xuanlin Li , Hao Su

In this paper, the problem of the trajectory design for a group of energy-constrained drones operating in dynamic wireless network environments is studied. In the considered model, a team of drone base stations (DBSs) is dispatched to…

Machine Learning · Computer Science 2020-12-08 Ye Hu , Mingzhe Chen , Walid Saad , H. Vincent Poor , Shuguang Cui

In 3D shape recognition, multi-view based methods leverage human's perspective to analyze 3D shapes and have achieved significant outcomes. Most existing research works in deep learning adopt handcrafted networks as backbones due to their…

Computer Vision and Pattern Recognition · Computer Science 2020-12-11 Zhaoqun Li , Hongren Wang , Jinxing Li

Classical pixel-based Visual Servoing (VS) approaches offer high accuracy but suffer from a limited convergence area due to optimization nonlinearity. Modern deep learning-based VS methods overcome traditional vision issues but lack…

Robotics · Computer Science 2023-10-03 Salar Asayesh , Hossein Sheikhi Darani , Mo chen , Mehran Mehrandezh , Kamal Gupta