English
Related papers

Related papers: SILVR: A Synthetic Immersive Large-Volume Plenopti…

200 papers

Recently, large vision-language models (LVLMs) unleash powerful analysis capabilities for low Earth orbit (LEO) satellite Earth observation images in the data center. However, fast satellite motion, brief satellite-ground station (GS)…

Networking and Internet Architecture · Computer Science 2025-07-09 Yuxin Zhang , Jiahao Yang , Zhe Chen , Wenjun Zhu , Jin Zhao , Yue Gao

We propose a novel framework for creating large-scale photorealistic datasets of indoor scenes, with ground truth geometry, material, lighting and semantics. Our goal is to make the dataset creation process widely accessible, transforming…

With dense inputs, Neural Radiance Fields (NeRF) is able to render photo-realistic novel views under static conditions. Although the synthesis quality is excellent, existing NeRF-based methods fail to obtain moderate three-dimensional (3D)…

Computer Vision and Pattern Recognition · Computer Science 2024-02-21 Shu Chen , Junyao Li , Yang Zhang , Beiji Zou

FVV Live is a novel end-to-end free-viewpoint video system, designed for low cost and real-time operation, based on off-the-shelf components. The system has been designed to yield high-quality free-viewpoint video using consumer-grade…

We present LiFMCR, a novel dataset for the registration of multiple micro lens array (MLA)-based light field cameras. While existing light field datasets are limited to single-camera setups and typically lack external ground truth, LiFMCR…

Computer Vision and Pattern Recognition · Computer Science 2025-10-16 Aymeric Fleith , Julian Zirbel , Daniel Cremers , Niclas Zeller

We introduce Plenoxels (plenoptic voxels), a system for photorealistic view synthesis. Plenoxels represent a scene as a sparse 3D grid with spherical harmonics. This representation can be optimized from calibrated images via gradient…

Computer Vision and Pattern Recognition · Computer Science 2021-12-10 Alex Yu , Sara Fridovich-Keil , Matthew Tancik , Qinhong Chen , Benjamin Recht , Angjoo Kanazawa

We present a new, publicly-available image dataset generated by the NVIDIA Deep Learning Data Synthesizer intended for use in object detection, pose estimation, and tracking applications. This dataset contains 144k stereo image pairs that…

Computer Vision and Pattern Recognition · Computer Science 2020-08-14 Mona Jalal , Josef Spjut , Ben Boudaoud , Margrit Betke

Manually tracking nutritional intake via food diaries is error-prone and burdensome. Automated computer vision techniques show promise for dietary monitoring but require large and diverse food image datasets. To address this need, we…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Saeejith Nair , Chi-en Amy Tai , Yuhao Chen , Alexander Wong

Plenoptic cameras use arrays of micro-lenses to capture multiple views of the same scene in a single compound image. They enable refocusing on different planes and depth estimation. However, until now, all types of plenoptic computational…

Optics · Physics 2020-01-29 K. M. Sowa , M. P. Kujda , P. Korecki

Photo-realistic free-viewpoint rendering of real-world scenes using classical computer graphics techniques is challenging, because it requires the difficult step of capturing detailed appearance and geometry models. Recent studies have…

Computer Vision and Pattern Recognition · Computer Science 2021-01-08 Lingjie Liu , Jiatao Gu , Kyaw Zaw Lin , Tat-Seng Chua , Christian Theobalt

Autonomous robots operating in natural karstic caves face perception and navigation challenges that are qualitatively distinct from those encountered in mines or tunnels: irregular geometry, reflective wet surfaces, near-zero ambient light,…

Restoring a sharp light field image from its blurry input has become essential due to the increasing popularity of parallax-based image processing. State-of-the-art blind light field deblurring methods suffer from several issues such as…

Computer Vision and Pattern Recognition · Computer Science 2019-12-06 Jonathan Samuel Lumentut , Tae Hyun Kim , Ravi Ramamoorthi , In Kyu Park

While Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in image and video understanding, their ability to comprehend the physical world has become an increasingly important research focus. Despite their…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Nanxi Li , Xiang Wang , Yuanjie Chen , Haode Zhang , Hong Li , Yong-Lu Li

Neural radiance fields (NeRFs) have become a ubiquitous tool for modeling scene appearance and geometry from multiview imagery. Recent work has also begun to explore how to use additional supervision from lidar or depth sensor measurements…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Anagh Malik , Parsa Mirdehghan , Sotiris Nousias , Kiriakos N. Kutulakos , David B. Lindell

First-Person-View (FPV) holds immense potential for revolutionizing the trajectory of Unmanned Aerial Vehicles (UAVs), offering an exhilarating avenue for navigating complex building structures. Yet, traditional Neural Radiance Field (NeRF)…

Computer Vision and Pattern Recognition · Computer Science 2024-08-13 Liqi Yan , Qifan Wang , Junhan Zhao , Qiang Guan , Zheng Tang , Jianhui Zhang , Dongfang Liu

We introduce OLATverse, a large-scale dataset comprising around 9M images of 765 real-world objects, captured from multiple viewpoints under a diverse set of precisely controlled lighting conditions. While recent advances in object-centric…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Xilong Zhou , Jianchun Chen , Pramod Rao , Timo Teufel , Linjie Lyu , Tigran Minasian , Oleksandr Sotnychenko , Xiao-Xiao Long , Marc Habermann , Christian Theobalt

The quest for deeper understanding of biological systems has driven the acquisition of increasingly larger multidimensional image datasets. Inspecting and manipulating data of this complexity is very challenging in traditional visualization…

Graphics · Computer Science 2018-08-23 Stanislav Pidhorskyi , Michael Morehead , Quinn Jones , George Spirou , Gianfranco Doretto

The operating room (OR) is an environment of interest for the development of sensing systems, enabling the detection of people, objects, and their semantic relations. Due to frequent occlusions in the OR, these systems often rely on input…

Computer Vision and Pattern Recognition · Computer Science 2023-08-31 Beerend G. A. Gerats , Jelmer M. Wolterink , Ivo A. M. J. Broeders

360 images represent scenes captured in all possible viewing directions and enable viewers to navigate freely around the scene thereby providing an immersive experience. Conversely, conventional images represent scenes in a single viewing…

Computer Vision and Pattern Recognition · Computer Science 2019-12-24 Julius Surya Sumantri , In Kyu Park

Scene-level novel view synthesis (NVS) is fundamental to many vision and graphics applications. Recently, pose-conditioned diffusion models have led to significant progress by extracting 3D information from 2D foundation models, but these…

Computer Vision and Pattern Recognition · Computer Science 2024-08-23 Joseph Tung , Gene Chou , Ruojin Cai , Guandao Yang , Kai Zhang , Gordon Wetzstein , Bharath Hariharan , Noah Snavely
‹ Prev 1 3 4 5 6 7 10 Next ›