English
Related papers

Related papers: MOISST: Multimodal Optimization of Implicit Scene …

200 papers

In this work, we present a new multi-view depth estimation method that utilizes both conventional reconstruction and learning-based priors over the recently proposed neural radiance fields (NeRF). Unlike existing neural network based…

Computer Vision and Pattern Recognition · Computer Science 2021-10-06 Yi Wei , Shaohui Liu , Yongming Rao , Wang Zhao , Jiwen Lu , Jie Zhou

Calibrating robots into their workspaces is crucial for manipulation tasks. Existing calibration techniques often rely on sensors external to the robot (cameras, laser scanners, etc.) or specialized tools. This reliance complicates the…

Robotics · Computer Science 2024-03-21 Podshara Chanrungmaneekul , Kejia Ren , Joshua T. Grace , Aaron M. Dollar , Kaiyu Hang

LiDAR-camera extrinsic calibration is essential for multi-modal data fusion in robotic perception systems. However, existing approaches typically rely on handcrafted calibration targets (e.g., checkerboards) or specific, static scene types,…

Robotics · Computer Science 2026-01-06 Zhiwei Huang , Yanwei Fu , Yi Zhou , Xieyuanli Chen , Qijun Chen , Rui Fan

Many real-world 3D reconstruction applications demand photorealism and metric accuracy across unbounded, complex scenes with challenging lighting and imperfect captures that current Neural Radiance Field (NeRF) pipelines only partly…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Vladislav Polianskii , Elijs Dima , Isabel Salmerón Marazuela , Gergő László Nagy , Sigurdur Sverrisson , Volodya Grancharov

Visually poor scenarios are one of the main sources of failure in visual localization systems in outdoor environments. To address this challenge, we present MOZARD, a multi-modal localization system for urban outdoor environments using…

Robotics · Computer Science 2020-03-04 Lukas Schaupp , Patrick Pfreundschuh , Mathias Buerki , Cesar Cadena , Roland Siegwart , Juan Nieto

This article describes a multi-modal method using simulated Lidar data via ray tracing and image pixel loss with differentiable rendering to optimize an object's position with respect to an observer or some referential objects in a computer…

Systems and Control · Electrical Eng. & Systems 2023-09-07 Sean Zanyk-McLean , Krishna Kumar , Paul Navratil

Cameras and LiDAR are essential sensors for autonomous vehicles. Camera-LiDAR data fusion compensate for deficiencies of stand-alone sensors but relies on precise extrinsic calibration. Many learning-based calibration methods predict…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Ni Ou , Zhuo Chen , Xinru Zhang , Junzheng Wang

Multi-sensor fusion using LiDAR and RGB cameras significantly enhances 3D object detection task. However, conventional LiDAR sensors perform dense, stateless scans, ignoring the strong temporal continuity in real-world scenes. This leads to…

Computer Vision and Pattern Recognition · Computer Science 2025-11-17 Sara Shoouri , Morteza Tavakoli Taba , Hun-Seok Kim

LiDAR-camera calibration is a precondition for many heterogeneous systems that fuse data from LiDAR and camera. However, the constraint from common field of view and the requirement for strict time synchronization make the calibration a…

Robotics · Computer Science 2019-07-31 Bo Fu , Yue Wang , Xiaqing Ding , Yanmei Jiao , Li Tang , Rong Xiong

The emerging Internet of Things (IoT) applications, such as driverless cars, have a growing demand for high-precision positioning and navigation. Nowadays, LiDAR inertial odometry becomes increasingly prevalent in robotics and autonomous…

Robotics · Computer Science 2025-03-10 Chengwei Zhao , Kun Hu , Jie Xu , Lijun Zhao , Baiwen Han , Kaidi Wu , Maoshan Tian , Shenghai Yuan

Sensor setups consisting of a combination of 3D range scanner lasers and stereo vision systems are becoming a popular choice for on-board perception systems in vehicles; however, the combined use of both sources of information implies a…

Computer Vision and Pattern Recognition · Computer Science 2017-07-28 Carlos Guindel , Jorge Beltrán , David Martín , Fernando García

While image registration has been studied in remote sensing community for decades, registering multimodal data [e.g., optical, LiDAR, SAR, and map] remains a challenging problem because of significant nonlinear intensity differences between…

Computer Vision and Pattern Recognition · Computer Science 2021-04-01 Yuanxin Ye , Lorenzo Bruzzone , Jie Shan , Francesca Bovolo , Qing Zhu

One of the current trends in robotics is to employ large language models (LLMs) to provide non-predefined command execution and natural human-robot interaction. It is useful to have an environment map together with its language…

Robotics · Computer Science 2025-01-09 Evgenii Kruzhkov , Sven Behnke

Multi-modal behaviors exhibited by surrounding vehicles (SVs) can typically lead to traffic congestion and reduce the travel efficiency of autonomous vehicles (AVs) in dense traffic. This paper proposes a real-time parallel trajectory…

Robotics · Computer Science 2023-09-12 Lei Zheng , Rui Yang , Zengqi Peng , Haichao Liu , Michael Yu Wang , Jun Ma

We aim to improve the Inverted Neural Radiance Fields (iNeRF) algorithm which defines the image pose estimation problem as a NeRF based iterative linear optimization. NeRFs are novel neural space representation models that can synthesize…

Computer Vision and Pattern Recognition · Computer Science 2023-10-06 Ágoston István Csehi , Csaba Máté Józsa

We present a method to accelerate global illumination computation in dynamic environments by taking advantage of limitations of the human visual system. A model of visual attention is used to locate regions of interest in a scene and to…

Graphics · Computer Science 2007-05-23 Yang Li Hector Yee

This paper proposes SemCal: an automatic, targetless, extrinsic calibration algorithm for a LiDAR and camera system using semantic information. We leverage a neural information estimator to estimate the mutual information (MI) of semantic…

Computer Vision and Pattern Recognition · Computer Science 2021-09-22 Peng Jiang , Philip Osteen , Srikanth Saripalli

Robust multisensor fusion of multi-modal measurements such as IMUs, wheel encoders, cameras, LiDARs, and GPS holds great potential due to its innate ability to improve resilience to sensor failures and measurement outliers, thereby enabling…

Robotics · Computer Science 2023-09-28 Woosik Lee , Patrick Geneva , Chuchu Chen , Guoquan Huang

Simultaneously odometry and mapping using LiDAR data is an important task for mobile systems to achieve full autonomy in large-scale environments. However, most existing LiDAR-based methods prioritize tracking quality over reconstruction…

Computer Vision and Pattern Recognition · Computer Science 2023-03-21 Junyuan Deng , Xieyuanli Chen , Songpengcheng Xia , Zhen Sun , Guoqing Liu , Wenxian Yu , Ling Pei

Multi-modal systems have the capacity of producing more reliable results than systems with a single modality in road detection due to perceiving different aspects of the scene. We focus on using raw sensor inputs instead of, as it is…

Robotics · Computer Science 2023-08-24 Erkan Milli , Özgür Erkent , Asım Egemen Yılmaz