中文
相关论文

相关论文: CleanMAP: Distilling Multimodal LLMs for Confidenc…

200 篇论文

Today's software stacks for autonomous vehicles rely on HD maps to enable sufficient localization, accurate path planning, and reliable motion prediction. Recent developments have resulted in pipelines for the automated generation of HD…

机器人学 · 计算机科学 2024-04-19 Maximilian Leitenstern , Florian Sauerbeck , Dominik Kulmer , Johannes Betz

While recent online HD mapping methods relieve burdened offline pipelines and solve map freshness, they remain limited by perceptual inaccuracies, occlusion in dense traffic, and an inability to fuse multi-agent observations. We propose…

计算机视觉与模式识别 · 计算机科学 2025-07-31 Yuheng Du , Sheng Yang , Lingxuan Wang , Zhenghua Hou , Chengying Cai , Zhitao Tan , Mingxia Chen , Shi-Sheng Huang , Qiang Li

High-definition (HD) map construction methods are crucial for providing precise and comprehensive static environmental information, which is essential for autonomous driving systems. While Camera-LiDAR fusion techniques have shown promising…

计算机视觉与模式识别 · 计算机科学 2025-07-03 Xiaoshuai Hao , Yuting Zhao , Yuheng Ji , Luanyuan Dai , Peng Hao , Dingzhe Li , Shuai Cheng , Rong Yin

Online high-definition (HD) map construction is an important and challenging task in autonomous driving. Recently, there has been a growing interest in cost-effective multi-view camera-based methods without relying on other sensors like…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Xiaoshuai Hao , Ruikai Li , Hui Zhang , Dingzhe Li , Rong Yin , Sangil Jung , Seung-In Park , ByungIn Yoo , Haimei Zhao , Jing Zhang

Currently, High-Definition (HD) maps are a prerequisite for the stable operation of autonomous vehicles. Such maps contain information about all static road objects for the vehicle to consider during navigation, such as road edges, road…

机器人学 · 计算机科学 2023-11-01 Mohamed Sayed , Stepan Perminov , Dzmitry Tsetserukou

Constructing HD semantic maps is a central component of autonomous driving. However, traditional pipelines require a vast amount of human efforts and resources in annotating and maintaining the semantics in the map, which limits its…

计算机视觉与模式识别 · 计算机科学 2022-03-21 Qi Li , Yue Wang , Yilun Wang , Hang Zhao

Multimodal large language models (MLLMs) have demonstrated significant progress in semantic scene understanding and text-image alignment, with reasoning variants enhancing performance on more complex tasks involving mathematics and logic.…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Sicheng Feng , Song Wang , Shuyi Ouyang , Lingdong Kong , Zikai Song , Jianke Zhu , Huan Wang , Xinchao Wang

Multimodal large language models (MLLMs) improve performance on vision-language tasks by integrating visual features from pre-trained vision encoders into large language models (LLMs). However, how MLLMs process and utilize visual…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Hao Yin , Guangzong Si , Zilei Wang

Online High-Definition (HD) map construction is pivotal for autonomous driving. While recent approaches leverage historical temporal fusion to improve performance, we identify a critical safety flaw in this paradigm: it is inherently…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Ruikai Li , Xinrun Li , Mengwei Xie , Hao Shan , Shoumeng Qiu , Xinyuan Chang , Yizhe Fan , Feng Xiong , Han Jiang , Yilong Ren , Haiyang Yu , Mu Xu , Yang Long , Varun Ojha , Zhiyong Cui

Current Large Language Models (LLMs) typically rely on coarse-grained national labels for pluralistic value alignment. However, such macro-level supervision often obscures intra-country value heterogeneity, yielding a loose alignment. We…

人工智能 · 计算机科学 2026-05-15 Pengyun Zhu , Yuqi Ren , Zhen Wang , Lei Yang , Deyi Xiong

Accurate online map matching is fundamental to vehicle navigation and the activation of intelligent driving functions. Current online map matching methods are prone to errors in complex road networks, especially in multilevel road area. To…

计算机视觉与模式识别 · 计算机科学 2025-05-13 Xin Bi , Zhichao Li , Yuxuan Xia , Panpan Tong , Lijuan Zhang , Yang Chen , Junsheng Fu

High-definition (HD) semantic map generation of the environment is an essential component of autonomous driving. Existing methods have achieved good performance in this task by fusing different sensor modalities, such as LiDAR and camera.…

计算机视觉与模式识别 · 计算机科学 2024-11-28 Hao Dong , Weihao Gu , Xianjing Zhang , Jintao Xu , Rui Ai , Huimin Lu , Juho Kannala , Xieyuanli Chen

Multimodal large language models (MLLMs) combine visual and textual data for tasks such as image captioning and visual question answering. Proper uncertainty calibration is crucial, yet challenging, for reliable use in areas like healthcare…

计算机视觉与模式识别 · 计算机科学 2024-12-30 Zijun Chen , Wenbo Hu , Guande He , Zhijie Deng , Zheng Zhang , Richang Hong

Crowdsourcing enables scalable autonomous driving map construction, but low-cost sensor noise hinders quality from improving with data volume. We propose CSMapping, a system that produces accurate semantic maps and topological road…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Zhijian Qiao , Zehuan Yu , Tong Li , Chih-Chung Chou , Wenchao Ding , Shaojie Shen

Large Vision-Language Models (LVLMs) have achieved impressive performance in multimodal tasks, but they still suffer from hallucinations, i.e., generating content that is grammatically accurate but inconsistent with visual inputs. In this…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Chenxi Li , Yichen Guo , Benfang Qian , Jinhao You , Kai Tang , Yaosong Du , Zonghao Zhang , Xiande Huang

Accurate environmental representations are essential for autonomous driving, providing the foundation for safe and efficient navigation. Traditionally, high-definition (HD) maps are providing this representation of the static road…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Thomas Monninger , Zihan Zhang , Steffen Staab , Sihao Ding

We propose a novel framework for filtering image-text data by leveraging fine-tuned Multimodal Language Models (MLMs). Our approach outperforms predominant filtering methods (e.g., CLIPScore) via integrating the recent advances in MLMs. We…

计算机视觉与模式识别 · 计算机科学 2024-03-06 Weizhi Wang , Khalil Mrini , Linjie Yang , Sateesh Kumar , Yu Tian , Xifeng Yan , Heng Wang

High-definition (HD) maps are essential for autonomous driving, yet multi-modal fusion often suffers from inconsistency between camera and LiDAR modalities, leading to performance degradation under low-light conditions, occlusions, or…

计算机视觉与模式识别 · 计算机科学 2026-02-26 Haoxiang Fu , Lingfeng Zhang , Hao Li , Ruibing Hu , Zhengrong Li , Guanjing Liu , Zimu Tan , Long Chen , Hangjun Ye , Xiaoshuai Hao

Online HD map construction is a fundamental task in autonomous driving systems, aiming to acquire semantic information of map elements around the ego vehicle based on real-time sensor inputs. Recently, several approaches have achieved…

计算机视觉与模式识别 · 计算机科学 2025-08-25 Ziyang Yan , Ruikai Li , Zhiyong Cui , Bohan Li , Han Jiang , Yilong Ren , Aoyong Li , Zhenning Li , Sijia Wen , Haiyang Yu

In recent years, the rapid development of high-precision map technology combined with artificial intelligence has ushered in a new development opportunity in the field of intelligent vehicles. High-precision map technology is an important…

人工智能 · 计算机科学 2024-02-27 Yong Wang , Yanlin Zhou , Huan Ji , Zheng He , Xinyu Shen
‹ 上一页 1 2 3 10 下一页 ›