中文
相关论文

相关论文: YawDD+: Frame-level Annotations for Accurate Yawn …

200 篇论文

Image labeling is a critical bottleneck in the development of computer vision technologies, often constraining machine learning performance due to the time-intensive nature of manual annotations. This work introduces a novel approach that…

计算机视觉与模式识别 · 计算机科学 2026-04-29 Amir Kazemi , Qurat ul ain Fatima , Volodymyr Kindratenko , Christopher W. Tessum

Deep learning architectures with powerful reasoning capabilities have driven significant advancements in autonomous driving technology. Large language models (LLMs) applied in this field can describe driving scenes and behaviors with a…

人工智能 · 计算机科学 2024-10-01 Yizhou Huang , Yihua Cheng , Kezhi Wang

Incorporating deep learning (DL) classification models into unmanned aerial vehicles (UAVs) can significantly augment search-and-rescue operations and disaster management efforts. In such critical situations, the UAV's ability to promptly…

图像与视频处理 · 电气工程与系统科学 2023-06-21 Gao Yu Lee , Tanmoy Dam , Md Meftahul Ferdaus , Daniel Puiu Poenar , Vu N. Duong

The key to ensuring the safe obstacle avoidance function of autonomous driving systems lies in the use of extremely accurate vehicle recognition techniques. However, the variability of the actual road environment and the diverse…

计算机视觉与模式识别 · 计算机科学 2025-01-03 Haocheng Guo , Yaqiong Zhang , Lieyang Chen , Arfat Ahmad Khan

Current deep learning paradigms largely benefit from the tremendous amount of annotated data. However, the quality of the annotations often varies among labelers. Multi-observer studies have been conducted to study these annotation…

计算机视觉与模式识别 · 计算机科学 2020-10-05 Xiaosong Wang , Ziyue Xu , Dong Yang , Leo Tam , Holger Roth , Daguang Xu

Detecting dangerous driving has been of critical interest for the past few years. However, a practical yet minimally intrusive solution remains challenging as existing technologies heavily rely on visual features or physical proximity. With…

人机交互 · 计算机科学 2023-09-22 Argha Sen , Avijit Mandal , Prasenjit Karmakar , Anirban Das , Sandip Chakraborty

Robotic manipulation and navigation are fundamental capabilities of embodied intelligence, enabling effective robot interactions with the physical world. Achieving these capabilities requires a cohesive understanding of the environment,…

机器人学 · 计算机科学 2025-11-18 Xiaoshuai Hao , Yingbo Tang , Lingfeng Zhang , Yanbiao Ma , Yunfeng Diao , Ziyu Jia , Wenbo Ding , Hangjun Ye , Long Chen

Spannotation is an open source user-friendly tool developed for image annotation for semantic segmentation specifically in autonomous navigation tasks. This study provides an evaluation of Spannotation, demonstrating its effectiveness in…

计算机视觉与模式识别 · 计算机科学 2024-02-29 Samuel O. Folorunsho , William R. Norris

A visually rich document (VRD) utilizes visual features along with linguistic cues to disseminate information. Training a custom extractor that identifies named entities from a document requires a large number of instances of the target…

Automated pavement distress detection via road images is still a challenging issue among pavement researchers and computer-vision community. In recent years, advancement in deep learning has enabled researchers to develop robust tools for…

机器学习 · 统计学 2020-04-29 Hamed Majidifard , Yaw Adu-Gyamfi , William G. Buttlar

Takeovers remain a key safety vulnerability in production ADAS, yet existing public resources rarely provide takeover-centered, real-world data. We present ADAS-TO, the first large-scale naturalistic dataset dedicated to ADAS-to-manual…

人机交互 · 计算机科学 2026-03-10 Yuhang Wang , Yiyao Xu , Jingran Sun , Hao Zhou

The recent emergence of Distributed Acoustic Sensing (DAS) technology has facilitated the effective capture of traffic-induced seismic data. The traffic-induced seismic wave is a prominent contributor to urban vibrations and contain crucial…

地球物理 · 物理学 2024-09-17 Xi Wang , Xin Liu , Songming Zhu , Zhanwen Li , Lina Gao

Response timing measures play a crucial role in the assessment of automated driving systems (ADS) in collision avoidance scenarios, including but not limited to establishing human benchmarks and comparing ADS to human driver response…

人机交互 · 计算机科学 2025-07-30 Shu-Yuan Liu , Johan Engström , Gustav Markkula

Autonomous driving is a popular research area within the computer vision research community. Since autonomous vehicles are highly safety-critical, ensuring robustness is essential for real-world deployment. While several public multimodal…

In the past several years, road anomaly segmentation is actively explored in the academia and drawing growing attention in the industry. The rationale behind is straightforward: if the autonomous car can brake before hitting an anomalous…

计算机视觉与模式识别 · 计算机科学 2024-01-11 Beiwen Tian , Huan-ang Gao , Leiyao Cui , Yupeng Zheng , Lan Luo , Baofeng Wang , Rong Zhi , Guyue Zhou , Hao Zhao

Accurate video annotation plays a vital role in modern retail applications, including customer behavior analysis, product interaction detection, and in-store activity recognition. However, conventional annotation methods heavily rely on…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Varun Mannam , Zhenyu Shi

Advanced Driver Assistance Systems (ADAS) have made driving safer over the last decade. They prepare vehicles for unsafe road conditions and alert drivers if they perform a dangerous maneuver. However, many accidents are unavoidable because…

机器人学 · 计算机科学 2016-01-06 Ashesh Jain , Hema S Koppula , Shane Soh , Bharad Raghavan , Avi Singh , Ashutosh Saxena

Current road damage detection methods, relying on manual inspections or sensor-mounted vehicles, are inefficient, limited in coverage, and often inaccurate, especially for minor damages, leading to delays and safety hazards. To address…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Weichao Pan , Jiaju Kang , Xu Wang , Zhihao Chen , Yiyuan Ge

The development of computer vision algorithms for Unmanned Aerial Vehicles (UAVs) imagery heavily relies on the availability of annotated high-resolution aerial data. However, the scarcity of large-scale real datasets with pixel-level…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Giulia Rizzoli , Francesco Barbato , Matteo Caligiuri , Pietro Zanuttigh

Point cloud data labeling is considered a time-consuming and expensive task in autonomous driving, whereas annotation-free learning training can avoid it by learning point cloud representations from unannotated data. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2025-01-08 Boyi Sun , Yuhang Liu , Xingxia Wang , Bin Tian , Long Chen , Fei-Yue Wang