中文
相关论文

相关论文: YOWOv3: An Efficient and Generalized Framework for…

200 篇论文

Learning from the limited amount of labeled data to the pre-train model has always been viewed as a challenging task. In this report, an effective and robust solution, the two-stage training paradigm YOLOv8 detector (TP-YOLOv8), is designed…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Zheng Wang , Dong Xie , Hanzhi Wang , Jiang Tian

With the accelerating pace of digital transformation and the widespread adoption of online platforms, both social and technical concerns regarding dark patterns-user interface designs that undermine users' ability to make informed and…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Se-Young Jang , Su-Yeon Yoon , Jae-Woong Jung , Dong-Hun Lee , Seong-Hun Choi , Soo-Kyung Jun , Yu-Bin Kim , Young-Seon Ju , Kyounggon Kim

Over the past years, YOLOs have emerged as the predominant paradigm in the field of real-time object detection owing to their effective balance between computational cost and detection performance. Researchers have explored the…

计算机视觉与模式识别 · 计算机科学 2024-10-31 Ao Wang , Hui Chen , Lihao Liu , Kai Chen , Zijia Lin , Jungong Han , Guiguang Ding

The increasing integration of sensors in autonomous maritime navigation has led to large-scale multimodal datasets, raising challenges in achieving efficient real-time perception. In such systems, object detection and trajectory perception…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Grigorios Papanikolaou , Ioannis Kontopoulos , Giannis Spiliopoulos , Dimitris Zissis , Konstantinos Tserpes

3D action recognition has broad applications in human-computer interaction and intelligent surveillance. However, recognizing similar actions remains challenging since previous literature fails to capture motion and shape cues effectively…

计算机视觉与模式识别 · 计算机科学 2017-12-08 Mengyuan Liu , Hong Liu , Chen Chen

Transformer-based human skeleton action recognition has been developed for years. However, the complexity and high parameter count demands of these models hinder their practical applications, especially in resource-constrained environments.…

计算机视觉与模式识别 · 计算机科学 2024-12-31 Wenhan Wu , Pengfei Wang , Chen Chen , Aidong Lu

The proposed YOLO-Former method seamlessly integrates the ideas of transformer and YOLOv4 to create a highly accurate and efficient object detection system. The method leverages the fast inference speed of YOLOv4 and incorporates the…

计算机视觉与模式识别 · 计算机科学 2024-01-15 Javad Khoramdel , Ahmad Moori , Yasamin Borhani , Armin Ghanbarzadeh , Esmaeil Najafi

Accurately detecting student behavior in classroom videos can aid in analyzing their classroom performance and improving teaching effectiveness. However, the current accuracy rate in behavior detection is low. To address this challenge, we…

计算机视觉与模式识别 · 计算机科学 2024-09-10 Fan Yang

This paper proposes a framework that combines online human state estimation, action recognition and motion prediction to enable early assessment and prevention of worker biomechanical risk during lifting tasks. The framework leverages the…

信号处理 · 电气工程与系统科学 2024-01-12 Cheng Guo , Lorenzo Rapetti , Kourosh Darvish , Riccardo Grieco , Francesco Draicchio , Daniele Pucci

The expanding applications, utilized by more users, enhance hardware performance and further develop cloud systems for big data processing. This leads to numerous unexplored deep learning applications, especially in advanced computer vision…

计算工程、金融与科学 · 计算机科学 2024-05-07 P. Veysi , M. Adeli , N. Peirov Naziri

Timely handgun detection is a crucial problem to improve public safety; nevertheless, the effectiveness of many surveillance systems still depends of finite human attention. Much of the previous research on handgun detection is based on…

计算机视觉与模式识别 · 计算机科学 2021-11-22 Mario Alberto Duran-Vega , Miguel Gonzalez-Mendoza , Leonardo Chang , Cuauhtemoc Daniel Suarez-Ramirez

Head detection provides distribution information of pedestrian, which is crucial for scene statistical analysis, traffic management, and risk assessment and early warning. However, scene complexity and large-scale variation in the real…

计算机视觉与模式识别 · 计算机科学 2023-10-17 Jiezhou Chen , Guankun Wang , Weixiang Liu , Xiaopin Zhong , Yibin Tian , ZongZe Wu

Latest CNN-based object detection models are quite accurate but require a high-performance GPU to run in real-time. They still are heavy in terms of memory size and speed for an embedded system with limited memory space. Since the object…

计算机视觉与模式识别 · 计算机科学 2021-08-03 Issac Sim , Ju-Hyung Lim , Young-Wan Jang , JiHwan You , SeonTaek Oh , Young-Keun Kim

This study examines the relationship between H.264 video compression and the performance of an object detection network (YOLOv5). We curated a set of 50 surveillance videos and annotated targets of interest (people, bikes, and vehicles).…

计算机视觉与模式识别 · 计算机科学 2022-11-14 Michael O'Byrne , Vibhoothi , Mark Sugrue , Anil Kokaram

Tiger conservation necessitates the strategic deployment of multifaceted initiatives encompassing the preservation of ecological habitats, anti-poaching measures, and community involvement for sustainable growth in the tiger population.…

计算机视觉与模式识别 · 计算机科学 2024-01-09 Gaurav Pendharkar , A. Ancy Micheal , Jason Misquitta , Ranjeesh Kaippada

This study presents a detailed analysis of the YOLOv8 object detection model, focusing on its architecture, training techniques, and performance improvements over previous iterations like YOLOv5. Key innovations, including the CSPNet…

计算机视觉与模式识别 · 计算机科学 2024-08-29 Muhammad Yaseen

Safe knife practices in the kitchen significantly reduce the risk of cuts, injuries, and serious accidents during food preparation. Using YOLOv7, an advanced object detection model, this study focuses on identifying safety risks during…

计算机视觉与模式识别 · 计算机科学 2025-01-10 Athulya Sundaresan Geetha

Recent graph convolutional neural networks (GCNs) have shown high performance in the field of human action recognition by using human skeleton poses. However, it fails to detect human-object interaction cases successfully due to the lack of…

计算机视觉与模式识别 · 计算机科学 2025-09-18 Hesham M. Shehata , Mohammad Abdolrahmani

Recent work on human animation usually incorporates large-scale video models, thereby achieving more vivid performance. However, the practical use of such methods is hindered by the slow inference speed and high computational demands.…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Rang Meng , Yan Wang , Weipeng Wu , Ruobing Zheng , Yuming Li , Chenguang Ma

A dominant paradigm for learning-based approaches in computer vision is training generic models, such as ResNet for image recognition, or I3D for video understanding, on large datasets and allowing them to discover the optimal…

计算机视觉与模式识别 · 计算机科学 2019-06-06 Yubo Zhang , Pavel Tokmakov , Martial Hebert , Cordelia Schmid