English
Related papers

Related papers: V2X-DGPE: Addressing Domain Gaps and Pose Errors f…

200 papers

Collaborative perception has the potential to significantly enhance perceptual accuracy through the sharing of complementary information among agents. However, real-world collaborative perception faces persistent challenges, particularly in…

Computer Vision and Pattern Recognition · Computer Science 2025-05-05 Zhengbin Zhang , Yan Wu , Hongkun Zhang

Due to domain shifts, machine learning systems typically struggle to generalize well to new domains that differ from those of training data, which is what domain generalization (DG) aims to address. Although a variety of DG methods have…

Machine Learning · Computer Science 2023-11-15 Jingang Qu , Thibault Faney , Ze Wang , Patrick Gallinari , Soleiman Yousef , Jean-Charles de Hemptinne

Vehicle-to-everything (V2X) autonomous driving opens up a promising direction for developing a new generation of intelligent transportation systems. Collaborative perception (CP) as an essential component to achieve V2X can overcome the…

Computer Vision and Pattern Recognition · Computer Science 2023-09-01 Si Liu , Chen Gao , Yuan Chen , Xingyu Peng , Xianghao Kong , Kun Wang , Runsheng Xu , Wentao Jiang , Hao Xiang , Jiaqi Ma , Miao Wang

We introduce a unified single and multi-view neural implicit 3D reconstruction framework VPFusion. VPFusion attains high-quality reconstruction using both - 3D feature volume to capture 3D-structure-aware context, and pixel-aligned image…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Jisan Mahmud , Jan-Michael Frahm

Camera-based 3D object detection and tracking are essential for perception in autonomous driving. Current state-of-the-art approaches often rely exclusively on either perspective-view (PV) or bird's-eye-view (BEV) features, limiting their…

Computer Vision and Pattern Recognition · Computer Science 2025-10-14 Markus Käppeler , Özgün Çiçek , Daniele Cattaneo , Claudius Gläser , Yakov Miron , Abhinav Valada

Roadside camera-driven 3D object detection is a crucial task in intelligent transportation systems, which extends the perception range beyond the limitations of vision-centric vehicles and enhances road safety. While previous studies have…

Computer Vision and Pattern Recognition · Computer Science 2024-09-17 Hao Shi , Chengshan Pang , Jiaming Zhang , Kailun Yang , Yuhao Wu , Huajian Ni , Yining Lin , Rainer Stiefelhagen , Kaiwei Wang

The capability to achieve high-precision positioning accuracy has been considered as one of the most critical requirements for vehicle-to-everything (V2X) services in the fifth-generation (5G) cellular networks. The non-line-of-sight (NLOS)…

Signal Processing · Electrical Eng. & Systems 2021-02-23 Abdurrahman Fouda , Ryan Keating , Amitava Ghosh

Domain generalization (DG) aims to learn a model on several source domains, hoping that the model can generalize well to unseen target domains. The distribution shift between domains contains the covariate shift and conditional shift, both…

Computer Vision and Pattern Recognition · Computer Science 2022-09-20 Jianxin Lin , Yongqiang Tang , Junping Wang , Wensheng Zhang

Due to the lack of depth cues in images, multi-frame inputs are important for the success of vision-based perception, prediction, and planning in autonomous driving. Observations from different angles enable the recovery of 3D object states…

Computer Vision and Pattern Recognition · Computer Science 2024-02-27 Yichen Xie , Hongge Chen , Gregory P. Meyer , Yong Jae Lee , Eric M. Wolff , Masayoshi Tomizuka , Wei Zhan , Yuning Chai , Xin Huang

Cooperative perception for connected and automated vehicles is traditionally achieved through the fusion of feature maps from two or more vehicles. However, the absence of feature maps shared from other vehicles can lead to a significant…

Computer Vision and Pattern Recognition · Computer Science 2024-08-28 Deyuan Qu , Qi Chen , Tianyu Bai , Hongsheng Lu , Heng Fan , Hao Zhang , Song Fu , Qing Yang

Promising complementarity exists between the texture features of color images and the geometric information of LiDAR point clouds. However, there still present many challenges for efficient and robust feature fusion in the field of 3D…

Computer Vision and Pattern Recognition · Computer Science 2022-09-16 Chaokang Jiang , Guangming Wang , Jinxing Wu , Yanzi Miao , Hesheng Wang

3D semantic occupancy prediction is an emerging perception paradigm in autonomous driving, providing a voxel-level representation of both geometric details and semantic categories. However, its effectiveness is inherently constrained in…

Computer Vision and Pattern Recognition · Computer Science 2026-01-19 Hanlin Wu , Pengfei Lin , Ehsan Javanmardi , Naren Bao , Bo Qian , Hao Si , Manabu Tsukada

Diffusion models are widely recognized for their ability to generate high-fidelity images. Despite the excellent performance and scalability of the Diffusion Transformer (DiT) architecture, it applies fixed compression across different…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Weinan Jia , Mengqi Huang , Nan Chen , Lei Zhang , Zhendong Mao

With a growing interest in autonomous vehicles' operation, there is an equally increasing need for efficient anticipatory gesture recognition systems for human-vehicle interaction. Existing gesture-recognition algorithms have been primarily…

Computer Vision and Pattern Recognition · Computer Science 2020-11-19 Nishant Bhattacharya , Suresh Sundaram

While RGBD-based methods for category-level object pose estimation hold promise, their reliance on depth data limits their applicability in diverse scenarios. In response, recent efforts have turned to RGB-based methods; however, they face…

Computer Vision and Pattern Recognition · Computer Science 2024-09-25 Ruida Zhang , Ziqin Huang , Gu Wang , Chenyangguang Zhang , Yan Di , Xingxing Zuo , Jiwen Tang , Xiangyang Ji

Fusing data from cameras and LiDAR sensors is an essential technique to achieve robust 3D object detection. One key challenge in camera-LiDAR fusion involves mitigating the large domain gap between the two sensors in terms of coordinates…

Computer Vision and Pattern Recognition · Computer Science 2023-02-17 Yecheol Kim , Konyul Park , Minwook Kim , Dongsuk Kum , Jun Won Choi

6D object pose estimation is widely applied in robotic tasks such as grasping and manipulation. Prior methods using RGB-only images are vulnerable to heavy occlusion and poor illumination, so it is important to complement them with depth…

Computer Vision and Pattern Recognition · Computer Science 2021-04-07 Yi Cheng , Hongyuan Zhu , Ying Sun , Cihan Acar , Wei Jing , Yan Wu , Liyuan Li , Cheston Tan , Joo-Hwee Lim

Collecting multi-view driving scenario videos to enhance the performance of 3D visual perception tasks presents significant challenges and incurs substantial costs, making generative models for realistic data an appealing alternative. Yet,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-29 Junpeng Jiang , Gangyi Hong , Miao Zhang , Hengtong Hu , Kun Zhan , Rui Shao , Liqiang Nie

Existing multi-agent perception algorithms usually select to share deep neural features extracted from raw sensing data between agents, achieving a trade-off between accuracy and communication bandwidth limit. However, these methods assume…

Computer Vision and Pattern Recognition · Computer Science 2023-03-14 Runsheng Xu , Jinlong Li , Xiaoyu Dong , Hongkai Yu , Jiaqi Ma

Recent cooperative perception datasets have played a crucial role in advancing smart mobility applications by enabling information exchange between intelligent agents, helping to overcome challenges such as occlusions and improving overall…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Karthikeyan Chandra Sekaran , Markus Geisler , Dominik Rößle , Adithya Mohan , Daniel Cremers , Wolfgang Utschick , Michael Botsch , Werner Huber , Torsten Schön