English
Related papers

Related papers: ClaraVid: A Holistic Scene Reconstruction Benchmar…

200 papers

In this thesis, we leverage monocular cameras on aerial robots to predict depth and semantic maps in low-altitude unstructured environments. We propose a joint deep-learning architecture, named Co-SemDepth, that can perform the two tasks…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Yara AlaaEldin

Single-view 3D reconstruction is currently approached from two dominant perspectives: reconstruction of scenes with limited diversity using 3D data supervision or reconstruction of diverse singular objects using large image priors. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Andreea Ardelean , Mert Özer , Bernhard Egger

The vision-based semantic scene completion task aims to predict dense geometric and semantic 3D scene representations from 2D images. However, the presence of dynamic objects in the scene seriously affects the accuracy of the model…

Computer Vision and Pattern Recognition · Computer Science 2025-05-06 Meng Wang , Fan Wu , Yunchuan Qin , Ruihui Li , Zhuo Tang , Kenli Li

We contribute a dense SLAM system that takes a live stream of depth images as input and reconstructs non-rigid deforming scenes in real time, without templates or prior models. In contrast to existing approaches, we do not maintain any…

Computer Vision and Pattern Recognition · Computer Science 2019-05-01 Wei Gao , Russ Tedrake

While foundation models drive steady progress in image segmentation and diffusion algorithms compose always more realistic images, the seemingly simple problem of identifying recurrent patterns in a collection of images remains very much…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Zeynep Sonat Baltacı , Romain Loiseau , Mathieu Aubry

Remarkable strides have been made in reconstructing static scenes or human bodies from monocular videos. Yet, the two problems have largely been approached independently, without much synergy. Most visual SLAM methods can only reconstruct…

Computer Vision and Pattern Recognition · Computer Science 2024-05-24 Yizhou Zhao , Tuanfeng Y. Wang , Bhiksha Raj , Min Xu , Jimei Yang , Chun-Hao Paul Huang

This work addresses the task of dense 3D reconstruction of a complex dynamic scene from images. The prevailing idea to solve this task is composed of a sequence of steps and is dependent on the success of several pipelines in its execution.…

Computer Vision and Pattern Recognition · Computer Science 2019-11-22 Suryansh Kumar , Yuchao Dai , Hongdong Li

Spatial pooling (SP) and cross-channel pooling (CCP) operators have been applied to aggregate spatial features and pixel-wise features from feature maps in deep neural networks (DNNs), respectively. Their main goal is to reduce computation…

Computer Vision and Pattern Recognition · Computer Science 2024-08-16 Xiaoqing Zhang , Qiushi Nie , Zunjie Xiao , Jilu Zhao , Xiao Wu , Pengxin Guo , Runzhi Li , Jin Liu , Yanjie Wei , Yi Pan

The advent of large aperture arrays, such as the ones currently under construction for the SKA project, allows for observing the Universe in the radio-spectrum at unprecedented resolution and sensitivity. To process the enormous amounts of…

Instrumentation and Methods for Astrophysics · Physics 2025-07-31 S. Wang , S. Mignot , S. Prunet , L. Di Mascolo , M. Spinelli , A. Ferrari

Timely assessment of structural damage is critical for disaster response and recovery. However, most prior work in natural disaster analysis relies on 2D imagery, which lacks depth, suffers from occlusions, and provides limited spatial…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Nhut Le , Ehsan Karimi , Maryam Rahnemoonfar

In the realm of autonomous driving, achieving precise 3D reconstruction of the driving environment is critical for ensuring safety and effective navigation. Neural Radiance Fields (NeRF) have shown promise in creating highly detailed and…

Computer Vision and Pattern Recognition · Computer Science 2024-07-26 Xiaochao Pan , Jiawei Yao , Hongrui Kou , Tong Wu , Canran Xiao

Piece-wise 3D planar reconstruction provides holistic scene understanding of man-made environments, especially for indoor scenarios. Most recent approaches focused on improving the segmentation and reconstruction results by introducing…

Computer Vision and Pattern Recognition · Computer Science 2022-02-01 Yaxu Xie , Fangwen Shu , Jason Rambach , Alain Pagani , Didier Stricker

Colloidoscope is a deep learning pipeline employing a 3D residual Unet architecture, designed to enhance the tracking of dense colloidal suspensions through confocal microscopy. This methodology uses a simulated training dataset that…

Dense matching is crucial for 3D scene reconstruction since it enables the recovery of scene 3D geometry from image acquisition. Deep Learning (DL)-based methods have shown effectiveness in the special case of epipolar stereo disparity…

Computer Vision and Pattern Recognition · Computer Science 2024-02-21 Teng Wu , Bruno Vallet , Marc Pierrot-Deseilligny , Ewelina Rupnik

Reconstructing building wireframe from airborne LiDAR point clouds yields a compact, topology-centric representation that enables structural understanding beyond dense meshes. Yet a key limitation persists: conventional methods have failed…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Donghyun Kim , Chanyoung Kim , Youngjoong Kwon , Seong Jae Hwang

Omnidirectional and 360{\deg} images are becoming widespread in industry and in consumer society, causing omnidirectional computer vision to gain attention. Their wide field of view allows the gathering of a great amount of information…

Computer Vision and Pattern Recognition · Computer Science 2024-01-31 Bruno Berenguel-Baeta , Jesus Bermudez-Cameo , Jose J. Guerrero

For low-level computer vision and image processing ML tasks, training on large datasets is critical for generalization. However, the standard practice of relying on real-world images primarily from the Internet comes with image quality,…

Computer Vision and Pattern Recognition · Computer Science 2022-12-09 Gyeongmin Choe , Beibei Du , Seonghyeon Nam , Xiaoyu Xiang , Bo Zhu , Rakesh Ranjan

Reconstructing 4D spatial intelligence from visual observations has long been a central yet challenging task in computer vision, with broad real-world applications. These range from entertainment domains like movies, where the focus is…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Yukang Cao , Jiahao Lu , Zhisheng Huang , Zhuowen Shen , Chengfeng Zhao , Fangzhou Hong , Zhaoxi Chen , Xin Li , Wenping Wang , Yuan Liu , Ziwei Liu

This paper addresses the challenge of reconstructing 3D indoor scenes from multi-view images. Many previous works have shown impressive reconstruction results on textured objects, but they still have difficulty in handling low-textured…

Computer Vision and Pattern Recognition · Computer Science 2022-05-19 Haoyu Guo , Sida Peng , Haotong Lin , Qianqian Wang , Guofeng Zhang , Hujun Bao , Xiaowei Zhou

Visual scene understanding is the core task in making any crucial decision in any computer vision system. Although popular computer vision datasets like Cityscapes, MS-COCO, PASCAL provide good benchmarks for several tasks (e.g. image…

Computer Vision and Pattern Recognition · Computer Science 2020-12-08 Maryam Rahnemoonfar , Tashnim Chowdhury , Argho Sarkar , Debvrat Varshney , Masoud Yari , Robin Murphy