English
Related papers

Related papers: SEF-MAP: Subspace-Decomposed Expert Fusion for Rob…

200 papers

The exploration of Bird's-Eye View (BEV) mapping technology has driven significant innovation in visual perception technology for autonomous driving. BEV mapping models need to be applied to the unlabeled real world, making the study of…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Siyu Li , Yihong Cao , Hao Shi , Yongsheng Zang , Xuan He , Kailun Yang , Zhiyong Li

Parameter-efficient fine-tuning (PEFT) techniques, such as prompts and adapters, are widely used in multi-modal tracking because they alleviate issues of full-model fine-tuning, including time inefficiency, high resource consumption,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Yabin Zhu , Jianqi Li , Chenglong Li , Jiaxiang Wang , Chengjie Gu , Jin Tang

Multimodal sensor fusion is an essential capability for autonomous robots, enabling object detection and decision-making in the presence of failing or uncertain inputs. While recent fusion methods excel in normal environmental conditions,…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Edoardo Palladin , Roland Dietze , Praveen Narayanan , Mario Bijelic , Felix Heide

The fusion of multimodal sensor streams, such as camera, lidar, and radar measurements, plays a critical role in object detection for autonomous vehicles, which base their decision making on these inputs. While existing methods exploit…

Computer Vision and Pattern Recognition · Computer Science 2020-07-01 Mario Bijelic , Tobias Gruber , Fahim Mannan , Florian Kraus , Werner Ritter , Klaus Dietmayer , Felix Heide

3D human pose estimation (HPE) in autonomous vehicles (AV) differs from other use cases in many factors, including the 3D resolution and range of data, absence of dense depth maps, failure modes for LiDAR, relative location between the…

Computer Vision and Pattern Recognition · Computer Science 2021-12-23 Jingxiao Zheng , Xinwei Shi , Alexander Gorban , Junhua Mao , Yang Song , Charles R. Qi , Ting Liu , Visesh Chari , Andre Cornman , Yin Zhou , Congcong Li , Dragomir Anguelov

Efficient data utilization is crucial for advancing 3D scene understanding in autonomous driving, where reliance on heavily human-annotated LiDAR point clouds challenges fully supervised methods. Addressing this, our study extends into…

Computer Vision and Pattern Recognition · Computer Science 2025-12-08 Lingdong Kong , Xiang Xu , Jiawei Ren , Wenwei Zhang , Liang Pan , Kai Chen , Wei Tsang Ooi , Ziwei Liu

Learning to reliably perceive and understand the scene is an integral enabler for robots to operate in the real-world. This problem is inherently challenging due to the multitude of object types as well as appearance changes caused by…

Computer Vision and Pattern Recognition · Computer Science 2021-11-05 Abhinav Valada , Rohit Mohan , Wolfram Burgard

Autonomous driving systems benefit from high-definition (HD) maps that provide critical information about road infrastructure. The online construction of HD maps offers a scalable approach to generate local maps from on-board sensors.…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Hongyu Lyu , Thomas Monninger , Julie Stephany Berrio Perez , Mao Shan , Zhenxing Ming , Stewart Worrall

We propose a novel scene-segmentation-based exposure compensation method for multi-exposure image fusion (MEF) based tone mapping. The aim of MEF-based tone mapping is to display high dynamic range (HDR) images on devices with limited…

Computer Vision and Pattern Recognition · Computer Science 2024-10-29 Yuma Kinoshita , Hitoshi Kiya

A recent sensor fusion in a Bird's Eye View (BEV) space has shown its utility in various tasks such as 3D detection, map segmentation, etc. However, the approach struggles with inaccurate camera BEV estimation, and a perception of distant…

Computer Vision and Pattern Recognition · Computer Science 2023-11-09 Minsu Kim , Giseop Kim , Kyong Hwan Jin , Sunwook Choi

Emotion recognition plays a vital role in enhancing human-computer interaction. In this study, we tackle the MER-SEMI challenge of the MER2025 competition by proposing a novel multimodal emotion recognition framework. To address the issue…

Computer Vision and Pattern Recognition · Computer Science 2025-08-11 Juewen Hu , Yexin Li , Jiulin Li , Shuo Chen , Pring Wong

Facial expression classification remains a challenging task due to the high dimensionality and inherent complexity of facial image data. This paper presents Hy-Facial, a hybrid feature extraction framework that integrates both deep learning…

Computer Vision and Pattern Recognition · Computer Science 2025-10-01 Xinjin Li , Yu Ma , Kaisen Ye , Jinghan Cao , Minghao Zhou , Yeyang Zhou

Semi-supervised learning addresses the issue of limited annotations in medical images effectively, but its performance is often inadequate for complex backgrounds and challenging tasks. Multi-modal fusion methods can significantly improve…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Dongdong Meng , Sheng Li , Hao Wu , Guoping Wang , Xueqing Yan

Most autonomous cars rely on the availability of high-definition (HD) maps. Current research aims to address this constraint by directly predicting HD map elements from onboard sensors and reasoning about the relationships between the…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Khanh Son Pham , Christian Witte , Jens Behley , Johannes Betz , Cyrill Stachniss

Recent advancements in high-definition \emph{HD} map construction have demonstrated the effectiveness of dense representations, which heavily rely on computationally intensive bird's-eye view \emph{BEV} features. While sparse…

Computer Vision and Pattern Recognition · Computer Science 2025-05-15 Anqing Jiang , Jinhao Chai , Yu Gao , Yiru Wang , Yuwen Heng , Zhigang Sun , Hao Sun , Zezhong Zhao , Li Sun , Jian Zhou , Lijuan Zhu , Shugong Xu , Hao Zhao

Lidars and cameras are critical sensors that provide complementary information for 3D detection in autonomous driving. While prevalent multi-modal methods simply decorate raw lidar point clouds with camera features and feed them directly to…

Computer Vision and Pattern Recognition · Computer Science 2022-03-17 Yingwei Li , Adams Wei Yu , Tianjian Meng , Ben Caine , Jiquan Ngiam , Daiyi Peng , Junyang Shen , Bo Wu , Yifeng Lu , Denny Zhou , Quoc V. Le , Alan Yuille , Mingxing Tan

Recent advances in high-definition (HD) map construction from surround-view images have highlighted their cost-effectiveness in deployment. However, prevailing techniques often fall short in accurately extracting and utilizing road…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Wenzhao Qiu , Shanmin Pang , Hao zhang , Jianwu Fang , Jianru Xue

Although multiview fusion has demonstrated potential in LiDAR segmentation, its dependence on computationally intensive point-based interactions, arising from the lack of fixed correspondences between views such as range view and Bird's-Eye…

Computer Vision and Pattern Recognition · Computer Science 2024-12-20 Shoumeng Qiu , Xinrun Li , XiangYang Xue , Jian Pu

Local map construction is a vital component of intelligent driving perception, offering necessary reference for vehicle positioning and planning. Standard Definition map (SDMap), known for its low cost, accessibility, and versatility, has…

Computer Vision and Pattern Recognition · Computer Science 2024-12-30 Jiaqi Li , Pingfan Jia , Jiaxing Chen , Jiaxi Liu , Lei He , Keqiang Li

LiDAR Semantic Segmentation is a fundamental task in autonomous driving perception consisting of associating each LiDAR point to a semantic label. Fully-supervised models have widely tackled this task, but they require labels for each scan,…

Computer Vision and Pattern Recognition · Computer Science 2024-11-06 Xavier Timoneda , Markus Herb , Fabian Duerr , Daniel Goehring , Fisher Yu
‹ Prev 1 3 4 5 6 7 10 Next ›