English
Related papers

Related papers: Cinematic Behavior Transfer via NeRF-based Differe…

200 papers

This project presents an exploration into 3D scene reconstruction of synthetic and real-world scenes using Neural Radiance Field (NeRF) approaches. We primarily take advantage of the reduction in training and rendering time of neural…

Computer Vision and Pattern Recognition · Computer Science 2022-10-25 Benedict Quartey , Tuluhan Akbulut , Wasiwasi Mgonzo , Zheng Xin Yong

With the widespread use of NeRF-based implicit 3D representation, the need for camera localization in the same representation becomes manifestly apparent. Doing so not only simplifies the localization process -- by avoiding an…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 Rashik Shrestha , Bishad Koju , Abhigyan Bhusal , Danda Pani Paudel , François Rameau

High resolution images can be acquired using a non-regular sampling sensor which consists of an underlying low resolution sensor that is covered with a non-regular sampling mask. The reconstructed high resolution image is then obtained…

Image and Video Processing · Electrical Eng. & Systems 2022-04-08 Markus Jonscher , Karina Jaskolka , Jürgen Seiler , André Kaup

Neural radiance fields (NeRFs) are a powerful tool for implicit scene representations, allowing for differentiable rendering and the ability to make predictions about unseen viewpoints. There has been growing interest in object and…

Robotics · Computer Science 2024-11-14 Boxuan Zhang , Lindsay Kleeman , Michael Burke

Collaborative mapping of unknown environments can be done faster and more robustly than a single robot. However, a collaborative approach requires a distributed paradigm to be scalable and deal with communication issues. This work presents…

Robotics · Computer Science 2025-08-08 Mahboubeh Asadi , Kourosh Zareinia , Sajad Saeedi

Visual simultaneous localization and mapping (SLAM) plays a critical role in autonomous robotic systems, especially where accurate and reliable measurements are essential for navigation and sensing. In feature-based SLAM, the quantityand…

Robotics · Computer Science 2025-09-03 Haolan Zhang , Chenghao Li , Thanh Nguyen Canh , Lijun Wang , Nak Young Chong

Asynchronously operating event cameras find many applications due to their high dynamic range, vanishingly low motion blur, low latency and low data bandwidth. The field saw remarkable progress during the last few years, and existing…

Computer Vision and Pattern Recognition · Computer Science 2023-03-27 Viktor Rudnev , Mohamed Elgharib , Christian Theobalt , Vladislav Golyanik

We cast multiview reconstruction from unknown pose as a generative modeling problem. From a collection of unannotated 2D images of a scene, our approach simultaneously learns both a network to predict camera pose from 2D image input, as…

Computer Vision and Pattern Recognition · Computer Science 2024-06-12 Xin Yuan , Rana Hanocka , Michael Maire

3D style transfer aims to generate stylized views of 3D scenes with specified styles, which requires high-quality generating and keeping multi-view consistency. Existing methods still suffer the challenges of high-quality stylization with…

Computer Vision and Pattern Recognition · Computer Science 2025-01-28 Zijiang Yang , Zhongwei Qiu , Chang Xu , Dongmei Fu

Accurate localization is essential for autonomous vehicles, yet sensor noise and drift over time can lead to significant pose estimation errors, particularly in long-horizon environments. A common strategy for correcting accumulated error…

Robotics · Computer Science 2025-12-18 Gaurav Bansal

The goal of our work is to generate high-quality novel views from monocular videos of complex and dynamic scenes. Prior methods, such as DynamicNeRF, have shown impressive performance by leveraging time-varying dynamic radiation fields.…

Computer Vision and Pattern Recognition · Computer Science 2024-07-03 Xingyu Miao , Yang Bai , Haoran Duan , Yawen Huang , Fan Wan , Yang Long , Yefeng Zheng

We present a method for automatically modifying a NeRF representation based on a single observation of a non-rigid transformed version of the original scene. Our method defines the transformation as a 3D flow, specifically as a weighted…

Computer Vision and Pattern Recognition · Computer Science 2024-06-18 Zhenggang Tang , Zhongzheng Ren , Xiaoming Zhao , Bowen Wen , Jonathan Tremblay , Stan Birchfield , Alexander Schwing

Modelling the mapping from scene irradiance to image intensity is essential for many computer vision tasks. Such mapping is known as the camera response. Most digital cameras use a nonlinear function to map irradiance, as measured by the…

Computer Vision and Pattern Recognition · Computer Science 2022-09-09 Yunfeng Zhao , Stuart Ferguson , Huiyu Zhou , Karen Rafferty

Recent trends in SLAM and visual navigation have embraced 3D Gaussians as the preferred scene representation, highlighting the importance of estimating camera poses from a single image using a pre-built Gaussian model. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-11-19 Hao Wang , Linqing Zhao , Xiuwei Xu , Jiwen Lu , Haibin Yan

Scene flow describes the motion of 3D objects in real world and potentially could be the basis of a good feature for 3D action recognition. However, its use for action recognition, especially in the context of convolutional neural networks…

Computer Vision and Pattern Recognition · Computer Science 2017-03-28 Pichao Wang , Wanqing Li , Zhimin Gao , Yuyao Zhang , Chang Tang , Philip Ogunbona

Recent advances in deep learning have significantly improved performance of video prediction. However, state-of-the-art methods still suffer from blurriness and distortions in their future predictions, especially when there are large…

Computer Vision and Pattern Recognition · Computer Science 2020-03-20 Osamu Shouno

In the proposed study, we describe an approach to improving the computational efficiency and robustness of visual SLAM algorithms on mobile robots with multiple cameras and limited computational power by implementing an intermediate layer…

We present CLIP-NeRF, a multi-modal 3D object manipulation method for neural radiance fields (NeRF). By leveraging the joint language-image embedding space of the recent Contrastive Language-Image Pre-Training (CLIP) model, we propose a…

Computer Vision and Pattern Recognition · Computer Science 2022-03-03 Can Wang , Menglei Chai , Mingming He , Dongdong Chen , Jing Liao

We present a framework, DISORF, to enable online 3D reconstruction and visualization of scenes captured by resource-constrained mobile robots and edge devices. To address the limited computing capabilities of edge devices and potentially…

Robotics · Computer Science 2024-08-05 Chunlin Li , Hanrui Fan , Xiaorui Huang , Ruofan Liang , Sankeerth Durvasula , Nandita Vijaykumar

When capturing images through the glass during rainy or snowy weather conditions, the resulting images often contain waterdrops adhered on the glass surface, and these waterdrops significantly degrade the image quality and performance of…

Computer Vision and Pattern Recognition · Computer Science 2024-04-01 Yunhao Li , Jing Wu , Lingzhe Zhao , Peidong Liu