English
Related papers

Related papers: SILVR: A Synthetic Immersive Large-Volume Plenopti…

200 papers

A light field image captures scenes through its micro-lens array, providing a rich representation that encompasses spatial and angular information. While this richness comes at significant data redundancy, most existing methods tend to…

Image and Video Processing · Electrical Eng. & Systems 2026-02-19 Zeke Zexi Hu , Haodong Chen , Hui Ye , Xiaoming Chen , Vera Yuk Ying Chung , Yiran Shen , Weidong Cai

Visual-inertial-odometry has attracted extensive attention in the field of autonomous driving and robotics. The size of Field of View (FoV) plays an important role in Visual-Odometry (VO) and Visual-Inertial-Odometry (VIO), as a large FoV…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Ze Wang , Kailun Yang , Hao Shi , Peng Li , Fei Gao , Kaiwei Wang

Currently almost all state-of-the-art novel view synthesis and reconstruction models rely on calibrated cameras or additional geometric priors for training. These prerequisites significantly limit their applicability to massive uncalibrated…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Ruoyu Wang , Yi Ma , Shenghua Gao

Image restoration methods like super-resolution and image synthesis have been successfully used in commercial cloud gaming products like NVIDIA's DLSS. However, restoration over gaming content is not well studied by the general public. The…

Computer Vision and Pattern Recognition · Computer Science 2024-09-02 Lebin Zhou , Kun Han , Nam Ling , Wei Wang , Wei Jiang

We present a conceptual framework for training Vision-Language Models (VLMs) to perform Visual Perspective Taking (VPT), a core capability for embodied cognition essential for Human-Robot Interaction (HRI). As a first step toward this goal,…

Artificial Intelligence · Computer Science 2025-05-21 Joel Currie , Gioele Migno , Enrico Piacenti , Maria Elena Giannaccini , Patric Bach , Davide De Tommaso , Agnieszka Wykowska

In this paper, we address the problem of simultaneous relighting and novel view synthesis of a complex scene from multi-view images with a limited number of light sources. We propose an analysis-synthesis approach called Relit-NeuLF.…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Zhong Li , Liangchen Song , Zhang Chen , Xiangyu Du , Lele Chen , Junsong Yuan , Yi Xu

Differentiable volumetric rendering is a powerful paradigm for 3D reconstruction and novel view synthesis. However, standard volume rendering approaches struggle with degenerate geometries in the case of limited viewpoint diversity, a…

Computer Vision and Pattern Recognition · Computer Science 2023-04-07 Vitor Guizilini , Igor Vasiljevic , Jiading Fang , Rares Ambrus , Sergey Zakharov , Vincent Sitzmann , Adrien Gaidon

Real-time 3D fluorescence microscopy is crucial for the spatiotemporal analysis of live organisms, such as neural activity monitoring. The eXtended field-of-view light field microscope (XLFM), also known as Fourier light field microscope,…

Image and Video Processing · Electrical Eng. & Systems 2023-06-16 Josué Page Vizcaíno , Panagiotis Symvoulidis , Zeguan Wang , Jonas Jelten , Paolo Favaro , Edward S. Boyden , Tobias Lasser

Advances in radiance fields have enabled photorealistic novel view synthesis. In several domains, large-scale real-world datasets have been developed to support comprehensive benchmarking and to facilitate progress beyond scene-specific…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Cheng-You Lu , Yi-Shan Hung , Wei-Ling Chi , Hao-Ping Wang , Charlie Li-Ting Tsai , Yu-Cheng Chang , Yu-Lun Liu , Thomas Do , Chin-Teng Lin

Semantic grids can be useful representations of the scene around an autonomous system. By having information about the layout of the space around itself, a robot can leverage this type of representation for crucial tasks such as navigation…

In this paper, we propose a Neural Radiance Fields (NeRF) based framework, referred to as Novel View Synthesis Framework (NVSF). It jointly learns the implicit neural representation of space and time-varying scene for both LiDAR and Camera.…

Computer Vision and Pattern Recognition · Computer Science 2025-06-26 Gaurav Sharma , Ravi Kothari , Josef Schmid

Light field presents a rich way to represent the 3D world by capturing the spatio-angular dimensions of the visual signal. However, the popular way of capturing light field (LF) via a plenoptic camera presents spatio-angular resolution…

Computer Vision and Pattern Recognition · Computer Science 2019-10-22 Anil Kumar Vadathya , Sharath Girish , Kaushik Mitra

The emerging 4D millimeter-wave radar, measuring the range, azimuth, elevation, and Doppler velocity of objects, is recognized for its cost-effectiveness and robustness in autonomous driving. Nevertheless, its point clouds exhibit…

Computer Vision and Pattern Recognition · Computer Science 2025-09-24 Yuzhi Wu , Li Xiao , Jun Liu , Guangfeng Jiang , XiangGen Xia

In this report, we summarize the first NTIRE challenge on light field (LF) image super-resolution (SR), which aims at super-resolving LF images under the standard bicubic degradation with a magnification factor of 4. This challenge develops…

Computer Vision and Pattern Recognition · Computer Science 2023-04-21 Yingqian Wang , Longguang Wang , Zhengyu Liang , Jungang Yang , Radu Timofte , Yulan Guo

We introduce a 3D-aware diffusion model, ZeroNVS, for single-image novel view synthesis for in-the-wild scenes. While existing methods are designed for single objects with masked backgrounds, we propose new techniques to address challenges…

Computer Vision and Pattern Recognition · Computer Science 2024-04-25 Kyle Sargent , Zizhang Li , Tanmay Shah , Charles Herrmann , Hong-Xing Yu , Yunzhi Zhang , Eric Ryan Chan , Dmitry Lagun , Li Fei-Fei , Deqing Sun , Jiajun Wu

Large vision-language models (LVLMs) achieve impressive performance, yet their internal decision-making processes remain opaque, making it difficult to determine if the success stems from true multimodal fusion or from reliance on unimodal…

Machine Learning · Computer Science 2026-04-01 Lixin Xiu , Xufang Luo , Hideki Nakayama

We propose DistillNeRF, a self-supervised learning framework addressing the challenge of understanding 3D environments from limited 2D observations in outdoor autonomous driving scenes. Our method is a generalizable feedforward model that…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Letian Wang , Seung Wook Kim , Jiawei Yang , Cunjun Yu , Boris Ivanovic , Steven L. Waslander , Yue Wang , Sanja Fidler , Marco Pavone , Peter Karkus

Training computer vision models usually requires collecting and labeling vast amounts of imagery under a diverse set of scene configurations and properties. This process is incredibly time-consuming, and it is challenging to ensure that the…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Yunhao Ge , Harkirat Behl , Jiashu Xu , Suriya Gunasekar , Neel Joshi , Yale Song , Xin Wang , Laurent Itti , Vibhav Vineet

We address the problem of synthesizing novel views from a monocular video depicting a complex dynamic scene. State-of-the-art methods based on temporally varying Neural Radiance Fields (aka dynamic NeRFs) have shown impressive results on…

Computer Vision and Pattern Recognition · Computer Science 2023-04-25 Zhengqi Li , Qianqian Wang , Forrester Cole , Richard Tucker , Noah Snavely

Scene understanding enables intelligent agents to interpret and comprehend their environment. While existing large vision-language models (LVLMs) for scene understanding have primarily focused on indoor household tasks, they face two…

Computer Vision and Pattern Recognition · Computer Science 2025-07-18 Penglei Sun , Yaoxian Song , Xiangru Zhu , Xiang Liu , Qiang Wang , Yue Liu , Changqun Xia , Tiefeng Li , Yang Yang , Xiaowen Chu