English
Related papers

Related papers: CP+: Camera Poses Augmentation with Large-scale Li…

200 papers

In this paper, we consider the color-plus-mono dual-camera system and propose an end-to-end convolutional neural network to align and fuse images from it in an efficient and cost-effective way. Our method takes cross-domain and cross-scale…

Computer Vision and Pattern Recognition · Computer Science 2022-09-08 Yaping Zhao , Haitian Zheng , Mengqi Ji , Ruqi Huang

Hundreds of millions of people routinely take photos using their smartphones as point and shoot (PAS) cameras, yet very few would have the photography skills to compose a good shot of a scene. While traditional PAS cameras have built-in…

Computer Vision and Pattern Recognition · Computer Science 2025-05-07 Jiawan Li , Fei Zhou , Zhipeng Zhong , Jiongzhi Lin , Guoping Qiu

Fusing 3D LiDAR features with 2D camera features is a promising technique for enhancing the accuracy of 3D detection, thanks to their complementary physical properties. While most of the existing methods focus on directly fusing camera…

Computer Vision and Pattern Recognition · Computer Science 2024-01-17 Lemeng Wu , Dilin Wang , Meng Li , Yunyang Xiong , Raghuraman Krishnamoorthi , Qiang Liu , Vikas Chandra

Co-Registration of aerial imagery and Light Detection and Ranging (LiDAR) data is quilt challenging because the different imaging mechanism causes significant geometric and radiometric distortions between such data. To tackle the problem,…

Computer Vision and Pattern Recognition · Computer Science 2020-04-22 Bai Zhu , Yuanxin Ye , Chao Yang , Liang Zhou , Huiyu Liu , Yungang Cao

In this work, we tackle the problem of category-level online pose tracking of objects from point cloud sequences. For the first time, we propose a unified framework that can handle 9DoF pose tracking for novel rigid object instances as well…

Computer Vision and Pattern Recognition · Computer Science 2021-10-22 Yijia Weng , He Wang , Qiang Zhou , Yuzhe Qin , Yueqi Duan , Qingnan Fan , Baoquan Chen , Hao Su , Leonidas J. Guibas

A 3D point cloud is often synthesized from depth measurements collected by sensors at different viewpoints. The acquired measurements are typically both coarse in precision and corrupted by noise. To improve quality, previous works denoise…

Image and Video Processing · Electrical Eng. & Systems 2020-02-12 Xue Zhang , Gene Cheung , Jiahao Pang , Dong Tian

Point cloud maps with accurate color are crucial in robotics and mapping applications. Existing approaches for producing RGB-colorized maps are primarily based on real-time localization using filter-based estimation or sliding window…

Robotics · Computer Science 2024-09-18 Rundong Li , Xiyuan Liu , Haotian Li , Zheng Liu , Jiarong Lin , Yixi Cai , Fu Zhang

Change detection and irregular object extraction in 3D point clouds is a challenging task that is of high importance not only for autonomous navigation but also for updating existing digital twin models of various industrial environments.…

Computer Vision and Pattern Recognition · Computer Science 2023-12-18 Nikolaos Stathoulopoulos , Anton Koval , George Nikolakopoulos

We present a fine-tuning method to improve the appearance of 3D geometries reconstructed from single images. We leverage advances in monocular depth estimation to obtain disparity maps and present a novel approach to transforming 2D…

Computer Vision and Pattern Recognition · Computer Science 2022-10-20 Marissa Ramirez de Chanlatte , Matheus Gadelha , Thibault Groueix , Radomir Mech

The demands on robotic manipulation skills to perform challenging tasks have drastically increased in recent times. To perform these tasks with dexterity, robots require perception tools to understand the scene and extract useful…

Robotics · Computer Science 2023-12-06 K. Samarawickrama , G. Sharma , A. Angleraud , R. Pieters

Video-based point cloud compression (V-PCC) converts the dynamic point cloud data into video sequences using traditional video codecs for efficient encoding. However, this lossy compression scheme introduces artifacts that degrade the color…

Computer Vision and Pattern Recognition · Computer Science 2024-12-20 Jingwei Bao , Yu Liu , Zeliang Li , Shuyuan Zhu , Siu-Kei Au Yeung

Rigid registration of multi-view and multi-platform LiDAR scans is a fundamental problem in 3D mapping, robotic navigation, and large-scale urban modeling applications. Data acquisition with LiDAR sensors involves scanning multiple areas…

Computer Vision and Pattern Recognition · Computer Science 2020-02-03 Aby Thomas , Adarsh Sunilkumar , Shankar Shylesh , Aby Abahai T. , Subhasree Methirumangalath , Dong Chen , Jiju Peethambaran

This paper presents a novel framework for robust 3D object detection from point clouds via cross-modal hallucination. Our proposed approach is agnostic to either hallucination direction between LiDAR and 4D radar. We introduce multiple…

Computer Vision and Pattern Recognition · Computer Science 2024-03-13 Jianning Deng , Gabriel Chan , Hantao Zhong , Chris Xiaoxuan Lu

LiDAR-camera systems have become increasingly popular in robotics recently. A critical and initial step in integrating the LiDAR and camera data is the calibration of the LiDAR-camera system. Most existing calibration methods rely on…

Robotics · Computer Science 2025-04-02 Shuyi Zhou , Shuxiang Xie , Ryoichi Ishikawa , Takeshi Oishi

For building a Augmented Reality (AR) pipeline, the most crucial step is the camera calibration as overall quality heavily depends on it. In turn camera calibration itself is influenced most by the choice of camera-to-pattern poses - yet…

Human-Computer Interaction · Computer Science 2019-07-10 Pavel Rojtberg

Lidars and cameras are critical sensors that provide complementary information for 3D detection in autonomous driving. While most prevalent methods progressively downscale the 3D point clouds and camera images and then fuse the high-level…

Computer Vision and Pattern Recognition · Computer Science 2023-09-22 Zixuan Yin , Han Sun , Ningzhong Liu , Huiyu Zhou , Jiaquan Shen

Point cloud segmentation (PCS) is to classify each point in point clouds. The task enables robots to parse their 3D surroundings and run autonomously. According to different point cloud representations, existing PCS models can be roughly…

Computer Vision and Pattern Recognition · Computer Science 2025-02-19 Bike Chen , Antti Tikanmäki , Juha Röning

Multi-camera systems offer rich observation capabilities for visual navigation and 3D scene reconstruction; however, the resulting feature redundancy often compromises computational efficiency. This challenge is particularly pronounced…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Shunkun Liang , Banglei Guan , Bin Li , Qifeng Yu , Yang Shang

LiDAR point clouds are widely used in autonomous driving and consist of large numbers of 3D points captured at high frequency to represent surrounding objects such as vehicles, pedestrians, and traffic signs. While this dense data enables…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Z. Rozsa , Á. Madaras , Q. Wei , X. Lu , M. Golarits , H. Yuan , T. Sziranyi , R. Hamzaoui

3D occupancy prediction aims to infer dense, voxel-wise scene semantics from sensor observations, where the 2D-to-3D view transformation serves as a crucial step in bridging image features and volumetric representations. Most previous…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Yuan Wu , Zhiqiang Yan , Jiawei Lian , Zhengxue Wang , Jian Yang