中文
相关论文

相关论文: Modeling and Measuring Redundancy in Multisource M…

200 篇论文

Detecting and tracking objects is a crucial component of any autonomous navigation method. For the past decades, object detection has yielded promising results using neural networks on various datasets. While many methods focus on…

计算机视觉与模式识别 · 计算机科学 2025-05-02 Mathis Morales , Golnaz Habibi

Driving scenes are inherently heterogeneous and dynamic. Multi-attribute scene identification, as a high-level visual perception capability, provides autonomous vehicles (AVs) with essential contextual awareness to understand, reason…

计算机视觉与模式识别 · 计算机科学 2025-09-19 Ke Li , Chenyu Zhang , Yuxin Ding , Xianbiao Hu , Ruwen Qin

Robust detection and tracking of objects is crucial for the deployment of autonomous vehicle technology. Image based benchmark datasets have driven development in computer vision tasks such as object detection, tracking and segmentation of…

The deployment of machine learning (ML)-based process monitoring systems has significantly advanced additive manufacturing (AM) by enabling real-time defect detection, quality assessment, and process optimization. However, redundancy is a…

计算工程、金融与科学 · 计算机科学 2025-05-01 Jiarui Xie , Yaoyao Fiona Zhao

Addressing missing modalities and limited labeled data is crucial for advancing robust multimodal learning. We propose Robult, a scalable framework designed to mitigate these challenges by preserving modality-specific information and…

机器学习 · 计算机科学 2025-09-26 Duy A. Nguyen , Abhi Kamboj , Minh N. Do

Diffusion models have significantly mitigated the impact of annotated data scarcity in remote sensing (RS). Although recent approaches have successfully harnessed these models to enable diverse and controllable Layout-to-Image (L2I)…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Xianbao Hou , Yonghao He , Zeyd Boukhers , John See , Hu Su , Wei Sui , Cong Yang

Autonomous vehicles (AVs) are expected to revolutionize transportation by improving efficiency and safety. Their success relies on 3D vision systems that effectively sense the environment and detect traffic agents. Among sensors AVs use to…

计算机视觉与模式识别 · 计算机科学 2025-09-24 Amirhesam Aghanouri , Cristina Olaverri-Monreal

High-definition (HD) map construction methods are crucial for providing precise and comprehensive static environmental information, which is essential for autonomous driving systems. While Camera-LiDAR fusion techniques have shown promising…

计算机视觉与模式识别 · 计算机科学 2025-07-03 Xiaoshuai Hao , Yuting Zhao , Yuheng Ji , Luanyuan Dai , Peng Hao , Dingzhe Li , Shuai Cheng , Rong Yin

Three-dimensional object detection is essential for autonomous driving and robotics, relying on effective fusion of multimodal data from cameras and radar. This work proposes RCDINO, a multimodal transformer-based model that enhances visual…

计算机视觉与模式识别 · 计算机科学 2025-08-22 Olga Matykina , Dmitry Yudin

Multi-modal 3D object detection models for automated driving have demonstrated exceptional performance on computer vision benchmarks like nuScenes. However, their reliance on densely sampled LiDAR point clouds and meticulously calibrated…

计算机视觉与模式识别 · 计算机科学 2024-04-23 Till Beemelmanns , Quan Zhang , Christian Geller , Lutz Eckstein

Vehicles provide an ideal platform for urban sensing applications, as they can be equipped with all kinds of sensing devices that can continuously monitor the environment around the travelling vehicle. In this work we are particularly…

网络与互联网体系结构 · 计算机科学 2021-09-24 R. Bruno , M. Nurchis

Trajectory prediction is crucial for the reliability and safety of autonomous driving systems, yet it remains a challenging task in complex interactive scenarios due to noisy trajectory observations and intricate agent interactions.…

计算机视觉与模式识别 · 计算机科学 2026-01-26 Wenyi Xiong , Jian Chen , Ziheng Qi , Wenhua Chen

The curation of large-scale datasets is still costly and requires much time and resources. Data is often manually labeled, and the challenge of creating high-quality datasets remains. In this work, we fill the research gap using active…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Ahmed Ghita , Bjørk Antoniussen , Walter Zimmer , Ross Greer , Christian Creß , Andreas Møgelmose , Mohan M. Trivedi , Alois C. Knoll

Remote sensing image object detection (RSIOD) aims to identify and locate specific objects within satellite or aerial imagery. However, there is a scarcity of labeled data in current RSIOD datasets, which significantly limits the…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Datao Tang , Xiangyong Cao , Xuan Wu , Jialin Li , Jing Yao , Xueru Bai , Dongsheng Jiang , Yin Li , Deyu Meng

Multimodal learning aims to improve performance by leveraging data from multiple sources. During joint multimodal training, due to modality bias, the advantaged modality often dominates backpropagation, leading to imbalanced optimization.…

机器学习 · 计算机科学 2025-11-19 Zhe Yang , Wenrui Li , Hongtao Chen , Penghong Wang , Ruiqin Xiong , Xiaopeng Fan

Large repositories of image-caption pairs are essential for the development of vision-language models. However, these datasets are often extracted from noisy data scraped from the web, and contain many mislabeled instances. In order to…

计算机视觉与模式识别 · 计算机科学 2025-06-06 Haoran Zhang , Aparna Balagopalan , Nassim Oufattole , Hyewon Jeong , Yan Wu , Jiacheng Zhu , Marzyeh Ghassemi

The driving environment perception has a vital role for autonomous driving and nowadays has been actively explored for its realization. The research community and relevant stakeholders necessitate the development of Deep Learning (DL)…

人工智能 · 计算机科学 2025-10-16 Jalal Khan , Manzoor Khan , Sherzod Turaev , Sumbal Malik , Hesham El-Sayed , Farman Ullah

Drivable areas and curbs are critical traffic elements for autonomous driving, forming essential components of the vehicle visual perception system and ensuring driving safety. Deep neural networks (DNNs) have significantly improved…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Fulong Ma , Daojie Peng , Jun Ma

The development of computer vision algorithms for Unmanned Aerial Vehicles (UAVs) imagery heavily relies on the availability of annotated high-resolution aerial data. However, the scarcity of large-scale real datasets with pixel-level…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Giulia Rizzoli , Francesco Barbato , Matteo Caligiuri , Pietro Zanuttigh

Large-scale datasets for single-label multi-class classification, such as \emph{ImageNet-1k}, have been instrumental in advancing deep learning and computer vision. However, a critical and often understudied aspect is the comprehensive…