English
Related papers

Related papers: ClaraVid: A Holistic Scene Reconstruction Benchmar…

200 papers

This work tackles 3D scene reconstruction for a video fly-over perspective problem in the maritime domain, with a specific emphasis on geometrically and visually sound reconstructions. This will allow for downstream tasks such as…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Rui Yi Yong , Samuel Picosson , Arnold Wiliem

Quantifying the gap between synthetic and real-world imagery is essential for improving both transformer-based models - that rely on large volumes of data - and datasets, especially in underexplored domains like aerial scene understanding…

Computer Vision and Pattern Recognition · Computer Science 2024-12-02 Alina Marcu

Stereo matching is a fundamental task for 3D scene reconstruction. Recently, deep learning based methods have proven effective on some benchmark datasets, such as KITTI and Scene Flow. UAVs (Unmanned Aerial Vehicles) are commonly utilized…

Computer Vision and Pattern Recognition · Computer Science 2023-02-21 Zhang Xiaoyi , Cao Xuefeng , Yu Anzhu , Yu Wenshuai , Li Zhenqi , Quan Yujun

Modern scene reconstruction methods are able to accurately recover 3D surfaces that are visible in one or more images. However, this leads to incomplete reconstructions, missing all occluded surfaces. While much progress has been made on…

Computer Vision and Pattern Recognition · Computer Science 2025-11-07 Sam Bahrami , Dylan Campbell

Significant progress has been made in photo-realistic scene reconstruction over recent years. Various disparate efforts have enabled capabilities such as multi-appearance or large-scale modeling; however, there lacks a welldesigned dataset…

Computer Vision and Pattern Recognition · Computer Science 2024-12-20 Xijun Liu , Yifan Zhou , Yuxiang Guo , Rama Chellappa , Cheng Peng

For many fundamental scene understanding tasks, it is difficult or impossible to obtain per-pixel ground truth labels from real images. We address this challenge by introducing Hypersim, a photorealistic synthetic dataset for holistic…

Computer Vision and Pattern Recognition · Computer Science 2021-08-19 Mike Roberts , Jason Ramapuram , Anurag Ranjan , Atulit Kumar , Miguel Angel Bautista , Nathan Paczan , Russ Webb , Joshua M. Susskind

We explore the task of geometric reconstruction of images captured from a mixture of ground and aerial views. Current state-of-the-art learning-based approaches fail to handle the extreme viewpoint variation between aerial-ground image…

Computer Vision and Pattern Recognition · Computer Science 2025-04-18 Khiem Vuong , Anurag Ghosh , Deva Ramanan , Srinivasa Narasimhan , Shubham Tulsiani

Reconstructing photo-realistic large-scale scenes from images, for example at city scale, is a long-standing problem in computer graphics. Neural rendering is an emerging technique that enables photo-realistic image synthesis from…

Graphics · Computer Science 2025-07-22 Yaru Liu , Derek Nowrouzezahri , Morgan Mcguire

In the context of Concentrated Solar Power (CSP) plants, aerial images captured by drones present a unique set of challenges. Unlike urban or natural landscapes commonly found in existing datasets, solar fields contain highly reflective…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 M. A. Pérez-Cutiño , J. Valverde , J. Capitán , J. M. Díaz-Báñez

Real-world aerial scene understanding is limited by a lack of datasets that contain densely annotated images curated under a diverse set of conditions. Due to inherent challenges in obtaining such images in controlled real-world settings,…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Sahil Khose , Anisha Pal , Aayushi Agarwal , Deepanshi , Judy Hoffman , Prithvijit Chattopadhyay

High Dynamic Range (HDR) content (i.e., images and videos) has a broad range of applications. However, capturing HDR content from real-world scenes is expensive and time-consuming. Therefore, the challenging task of reconstructing visually…

Computer Vision and Pattern Recognition · Computer Science 2024-03-28 Hrishav Bakul Barua , Kalin Stefanov , KokSheik Wong , Abhinav Dhall , Ganesh Krishnasamy

In this work, we use multi-view aerial images to reconstruct the geometry, lighting, and material of facades using neural signed distance fields (SDFs). Without the requirement of complex equipment, our method only takes simple RGB images…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Zixuan Xie , Rengan Xie , Rong Li , Kai Huang , Pengju Qiao , Jingsen Zhu , Xu Yin , Qi Ye , Wei Hua , Yuchi Huo , Hujun Bao

We have witnessed significant progress in deep learning-based 3D vision, ranging from neural radiance field (NeRF) based 3D representation learning to applications in novel view synthesis (NVS). However, existing scene-level datasets for…

The state of the art in human-centric computer vision achieves high accuracy and robustness across a diverse range of tasks. The most effective models in this domain have billions of parameters, thus requiring extremely large datasets,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-22 Fatemeh Saleh , Sadegh Aliakbarian , Charlie Hewitt , Lohit Petikam , Xiao-Xian , Antonio Criminisi , Thomas J. Cashman , Tadas Baltrušaitis

Deep learning has largely reshaped remote sensing (RS) research for aerial image understanding and made a great success. Nevertheless, most of the existing deep models are initialized with the ImageNet pretrained weights. Since natural…

Computer Vision and Pattern Recognition · Computer Science 2023-07-19 Di Wang , Jing Zhang , Bo Du , Gui-Song Xia , Dacheng Tao

Differentiable volumetric rendering is a powerful paradigm for 3D reconstruction and novel view synthesis. However, standard volume rendering approaches struggle with degenerate geometries in the case of limited viewpoint diversity, a…

Computer Vision and Pattern Recognition · Computer Science 2023-04-07 Vitor Guizilini , Igor Vasiljevic , Jiading Fang , Rares Ambrus , Sergey Zakharov , Vincent Sitzmann , Adrien Gaidon

Recent progress of deep image classification models has provided great potential to improve state-of-the-art performance in related computer vision tasks. However, the transition to semantic segmentation is hampered by strict memory…

Computer Vision and Pattern Recognition · Computer Science 2019-05-15 Ivan Krešo , Josip Krapac , Siniša Šegvić

Sparse and feature SLAM methods provide robust camera pose estimation. However, they often fail to capture the level of detail required for inspection and scene awareness tasks. Conversely, dense SLAM approaches generate richer scene…

Robotics · Computer Science 2025-05-16 Maaz Qureshi , Alexander Werner , Zhenan Liu , Amir Khajepour , George Shaker , William Melek

There has been a recent surge in methods that aim to decompose and segment scenes into multiple objects in an unsupervised manner, i.e., unsupervised multi-object segmentation. Performing such a task is a long-standing goal of computer…

Computer Vision and Pattern Recognition · Computer Science 2021-11-22 Laurynas Karazija , Iro Laina , Christian Rupprecht

Open-world 3D scene understanding is a critical challenge that involves recognizing and distinguishing diverse objects and categories from 3D data, such as point clouds, without relying on manual annotations. Traditional methods struggle…

Computer Vision and Pattern Recognition · Computer Science 2025-09-18 Yuru Wang , Pei Liu , Songtao Wang , Zehan Zhang , Xinyan Lu , Changwei Cai , Hao Li , Fu Liu , Peng Jia , Xianpeng Lang
‹ Prev 1 2 3 10 Next ›