中文
相关论文

相关论文: YOLOatr : Deep Learning Based Automatic Target Det…

200 篇论文

In high-risk railway construction, personal protective equipment monitoring is critical but challenging due to small and frequently obstructed targets. We propose YOLO-EA, an innovative model that enhances safety measure detection by…

计算机视觉与模式识别 · 计算机科学 2024-11-06 Hao Liu , Xue Qin

The You Only Look Once (YOLO) architecture is crucial for real-time object detection. However, deploying it in resource-constrained environments such as unmanned aerial vehicles (UAVs) requires efficient transfer learning. Although layer…

计算机视觉与模式识别 · 计算机科学 2025-09-09 Andrzej D. Dobrzycki , Ana M. Bernardos , José R. Casar

Nowadays, plenty of deep learning technologies are being applied to all aspects of autonomous driving with promising results. Among them, object detection is the key to improve the ability of an autonomous agent to perceive its environment…

计算机视觉与模式识别 · 计算机科学 2021-08-02 Yongxiang Gu , Qianlei Wang , Xiaolin Qin

Modern object detectors are static, fixed-depth networks optimized for a single operating point, requiring separate models for different deployment scenarios. We present an any-depth detection framework that enables a single network to span…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Woochul Kang , Hyungseop Lee , Jiho Lee

In this paper, we aim to improve the performance of a deep learning model towards image classification tasks, proposing a novel anchor-based training methodology, named \textit{Online Anchor-based Training} (OAT). The OAT method, guided by…

计算机视觉与模式识别 · 计算机科学 2024-06-19 Maria Tzelepi , Vasileios Mezaris

Image-Text Retrieval (ITR) finds broad applications in healthcare, aiding clinicians and radiologists by automatically retrieving relevant patient cases in the database given the query image and/or report, for more efficient clinical…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Meng Zheng , Jiajin Zhang , Benjamin Planche , Zhongpai Gao , Terrence Chen , Ziyan Wu

This study examines the effectiveness of spatio-temporal modeling and the integration of spatial attention mechanisms in deep learning models for underwater object detection. Specifically, in the first phase, the performance of…

计算机视觉与模式识别 · 计算机科学 2025-10-31 Sai Likhith Karri , Ansh Saxena

Multi-object tracking (MOT) is a vital component of intelligent video analytics applications such as surveillance and autonomous driving. The time and storage complexity required to execute deep learning models for visual object tracking…

计算机视觉与模式识别 · 计算机科学 2022-10-28 Keivan Nalaie , Rong Zheng

Manual identification of archaeological features in LiDAR imagery is labor-intensive, costly, and requires archaeological expertise. This paper shows how recent advancements in deep learning (DL) present efficient solutions for accurately…

计算机视觉与模式识别 · 计算机科学 2024-03-12 Jincheng Zhang , William Ringle , Andrew R. Willis

This paper presents a comprehensive solution to address the critical challenge of liquid leaks in the oil and gas industry, leveraging advanced computer vision and deep learning methodologies. Employing You Only Look Once (YOLO) and…

计算机视觉与模式识别 · 计算机科学 2023-12-19 Kalpak Bansod , Yanshan Wan , Yugesh Rai

For the autonomous drone-based inspection of wind turbine (WT) blades, accurate detection of the WT and its key features is essential for safe drone positioning and collision avoidance. Existing deep learning methods typically rely on…

计算机视觉与模式识别 · 计算机科学 2025-07-30 Arash Shahirpour , Jakob Gebler , Manuel Sanders , Tim Reuscher

Pedestrians and bicyclists are among the vulnerable road users (VRUs) that are inherently exposed to intricate traffic scenarios, which puts them at increased risk of sustaining injuries or facing fatal outcomes. This study presents an…

图像与视频处理 · 电气工程与系统科学 2025-07-16 Faryal Aurooj Nasir , Salman Liaquat , Nor Muzlifah Mahyuddin

This study presents an Adaptive Transfer Learning and Thresholding-based Deep Learning Model (ATL-TDLM) for automated breathing pattern recognition using thermal imaging. Unlike conventional methods that rely on sound-based respiratory…

图像与视频处理 · 电气工程与系统科学 2026-04-21 Hamza Kheddar , Yassine Himeur , Abbes Amira

This paper presents a robust approach for object detection in aerial imagery using the YOLOv5 model. We focus on identifying critical objects such as ambulances, car crashes, police vehicles, tow trucks, fire engines, overturned cars, and…

计算机视觉与模式识别 · 计算机科学 2025-01-09 Sindhu Boddu , Arindam Mukherjee

Improving the accuracy of fire detection using infrared night vision cameras remains a challenging task. Previous studies have reported strong performance with popular detection models. For example, YOLOv7 achieved an mAP50-95 of 0.51 using…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Nguyen Truong Khai , Luong Duc Vinh

Autonomous unmanned aerial vehicles (UAVs) integrated with edge computing capabilities empower real-time data processing directly on the device, dramatically reducing latency in critical scenarios such as wildfire detection. This study…

计算机视觉与模式识别 · 计算机科学 2025-01-16 Giovanny Vazquez , Shengjie Zhai , Mei Yang

We introduce, XoFTR, a cross-modal cross-view method for local feature matching between thermal infrared (TIR) and visible images. Unlike visible images, TIR images are less susceptible to adverse lighting and weather conditions but present…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Önder Tuzcuoğlu , Aybora Köksal , Buğra Sofu , Sinan Kalkan , A. Aydın Alatan

Rapid advances in deep learning for computer vision have driven the adoption of RGB camera-based adaptive traffic light systems to improve traffic safety and pedestrian comfort. However, these systems often overlook the needs of people with…

计算机视觉与模式识别 · 计算机科学 2025-05-15 Xiao Ni , Carsten Kuehnel , Xiaoyi Jiang

Along with the improvement of radar technologies, Automatic Target Recognition (ATR) using Synthetic Aperture Radar (SAR) and Inverse SAR (ISAR) has come to be an active research area. SAR/ISAR are radar techniques to generate a…

计算机视觉与模式识别 · 计算机科学 2018-03-13 Carlos Pena-Caballero , Elifaleth Cantu , Jesus Rodriguez , Adolfo Gonzales , Osvaldo Castellanos , Angel Cantu , Megan Strait , Jae Son , Dongchul Kim

This paper explores the application of knowledge distillation technology in target detection tasks, especially the impact of different distillation temperatures on the performance of student models. By using YOLOv5l as the teacher network…

计算机视觉与模式识别 · 计算机科学 2024-10-17 Guanming Huang , Aoran Shen , Yuxiang Hu , Junliang Du , Jiacheng Hu , Yingbin Liang