中文
相关论文

相关论文: MEVDT: Multi-Modal Event-Based Vehicle Detection a…

200 篇论文

We then introduce a novel hierarchical knowledge distillation strategy that incorporates the similarity matrix, feature representation, and response map-based distillation to guide the learning of the student Transformer network. We also…

计算机视觉与模式识别 · 计算机科学 2025-02-11 Shiao Wang , Xiao Wang , Chao Wang , Liye Jin , Lin Zhu , Bo Jiang , Yonghong Tian , Jin Tang

Incorporating multiple camera views for detection alleviates the impact of occlusions in crowded scenes. In a multiview system, we need to answer two important questions when dealing with ambiguities that arise from occlusions. First, how…

计算机视觉与模式识别 · 计算机科学 2021-05-04 Yunzhong Hou , Liang Zheng , Stephen Gould

Achieving level-5 driving automation in autonomous vehicles necessitates a robust semantic visual perception system capable of parsing data from different sensors across diverse conditions. However, existing semantic perception datasets…

计算机视觉与模式识别 · 计算机科学 2025-01-28 Tim Brödermann , David Bruggemann , Christos Sakaridis , Kevin Ta , Odysseas Liagouris , Jason Corkill , Luc Van Gool

Multi-view object tracking (MVOT) offers promising solutions to challenges such as occlusion and target loss, which are common in traditional single-view tracking. However, progress has been limited by the lack of comprehensive multi-view…

计算机视觉与模式识别 · 计算机科学 2025-02-28 Mengjie Xu , Yitao Zhu , Haotian Jiang , Jiaming Li , Zhenrong Shen , Sheng Wang , Haolin Huang , Xinyu Wang , Qing Yang , Han Zhang , Qian Wang

Object detection in event streams has emerged as a cutting-edge research area, demonstrating superior performance in low-light conditions, scenarios with motion blur, and rapid movements. Current detectors leverage spiking neural networks,…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Xiao Wang , Yu Jin , Wentao Wu , Wei Zhang , Lin Zhu , Bo Jiang , Yonghong Tian

Perceiving the surrounding environment is a fundamental task in autonomous driving. To obtain highly accurate perception results, modern autonomous driving systems typically employ multi-modal sensors to collect comprehensive environmental…

计算机视觉与模式识别 · 计算机科学 2024-09-10 Zhiwei Lin , Zhe Liu , Yongtao Wang , Le Zhang , Ce Zhu

Accurate 3D object detection for autonomous driving requires complementary sensors. Cameras provide dense semantics but unreliable depth, while millimeter-wave radar offers precise range and velocity measurements with sparse geometry. We…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Mayank Mayank , Bharanidhar Duraisamy , Florian Geiß , Abhinav Valada

As an alternative sensing paradigm, dynamic vision sensors (DVS) have been recently explored to tackle scenarios where conventional sensors result in high data rate and processing time. This paper presents a hybrid event-frame approach for…

计算机视觉与模式识别 · 计算机科学 2022-05-11 Vivek Mohan , Deepak Singla , Tarun Pulluri , Andres Ussa , Pradeep Kumar Gopalakrishnan , Pao-Sheng Sun , Bharath Ramesh , Arindam Basu

We present MVMO (Multi-View, Multi-Object dataset): a synthetic dataset of 116,000 scenes containing randomly placed objects of 10 distinct classes and captured from 25 camera locations in the upper hemisphere. MVMO comprises…

计算机视觉与模式识别 · 计算机科学 2022-06-01 Aitor Alvarez-Gila , Joost van de Weijer , Yaxing Wang , Estibaliz Garrote

One of the challenges in evaluating multi-object video detection, tracking and classification systems is having publically available data sets with which to compare different systems. However, the measures of performance for tracking and…

计算机视觉与模式识别 · 计算机科学 2017-04-24 Avishek Chakraborty , Victor Stamatescu , Sebastien C. Wong , Grant Wigley , David Kearney

This paper introduces the Emirates Multi-Task (EMT) dataset, designed to support multi-task benchmarking within a unified framework. It comprises over 30,000 frames from a dash-camera perspective and 570,000 annotated bounding boxes,…

计算机视觉与模式识别 · 计算机科学 2025-05-26 Nadya Abdel Madjid , Murad Mebrahtu , Abdulrahman Ahmad , Abdelmoamen Nasser , Bilal Hassan , Naoufel Werghi , Jorge Dias , Majid Khonji

In this paper, we present DSERT-RoLL, a driving dataset that incorporates stereo event, RGB, and thermal cameras together with 4D radar and dual LiDAR, collected across diverse weather and illumination conditions. The dataset provides…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Hoonhee Cho , Jae-Young Kang , Yuhwan Jeong , Yunseo Yang , Wonyoung Lee , Youngho Kim , Kuk-Jin Yoon

Event-based cameras, also called silicon retinas, potentially revolutionize computer vision by detecting and reporting significant changes in intensity asynchronous events, offering extended dynamic range, low latency, and low power…

图像与视频处理 · 电气工程与系统科学 2023-11-06 Julian Moosmann , Jakub Mandula , Philipp Mayer , Luca Benini , Michele Magno

General-domain large multimodal models (LMMs) have achieved significant advances in various image-text tasks. However, their performance in the Intelligent Traffic Surveillance (ITS) domain remains limited due to the absence of dedicated…

计算机视觉与模式识别 · 计算机科学 2025-09-15 Kaikai Zhao , Zhaoxiang Liu , Peng Wang , Xin Wang , Zhicheng Ma , Yajun Xu , Wenjing Zhang , Yibing Nan , Kai Wang , Shiguo Lian

We propose an online tracking algorithm that performs the object detection and data association under a common framework, capable of linking objects after a long time span. This is realized by preserving a large spatio-temporal memory to…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Jiarui Cai , Mingze Xu , Wei Li , Yuanjun Xiong , Wei Xia , Zhuowen Tu , Stefano Soatto

Agile locomotion in legged robots poses significant challenges for visual perception. Traditional frame-based cameras often fail in these scenarios for producing blurred images, particularly under low-light conditions. In contrast, event…

机器人学 · 计算机科学 2026-01-07 Jingcheng Cao , Chaoran Xiong , Jianmin Song , Shang Yan , Jiachen Liu , Ling Pei

Cooperative perception offers several benefits for enhancing the capabilities of autonomous vehicles and improving road safety. Using roadside sensors in addition to onboard sensors increases reliability and extends the sensor range.…

计算机视觉与模式识别 · 计算机科学 2024-03-05 Walter Zimmer , Gerhard Arya Wardana , Suren Sritharan , Xingcheng Zhou , Rui Song , Alois C. Knoll

Events cameras are ideal sensors for enabling robots to detect and track objects in highly dynamic environments due to their low latency output, high temporal resolution, and high dynamic range. In this paper, we present the Asynchronous…

计算机视觉与模式识别 · 计算机科学 2025-12-15 Angus Apps , Ziwei Wang , Vladimir Perejogin , Timothy Molloy , Robert Mahony

Event cameras, such as dynamic vision sensors (DVS), and dynamic and active-pixel vision sensors (DAVIS) can supplement other autonomous driving sensors by providing a concurrent stream of standard active pixel sensor (APS) images and DVS…

计算机视觉与模式识别 · 计算机科学 2017-11-07 Jonathan Binas , Daniel Neil , Shih-Chii Liu , Tobi Delbruck

Vehicle-to-Vehicle (V2V) cooperative perception has great potential to enhance autonomous driving performance by overcoming perception limitations in complex adverse traffic scenarios (CATS). Meanwhile, data serves as the fundamental…