中文
相关论文

相关论文: Vi-SAFE: A Spatial-Temporal Framework for Efficien…

200 篇论文

The proposed YOLO-Former method seamlessly integrates the ideas of transformer and YOLOv4 to create a highly accurate and efficient object detection system. The method leverages the fast inference speed of YOLOv4 and incorporates the…

计算机视觉与模式识别 · 计算机科学 2024-01-15 Javad Khoramdel , Ahmad Moori , Yasamin Borhani , Armin Ghanbarzadeh , Esmaeil Najafi

Recent text-to-video (T2V) models can synthesize complex videos from lightweight natural language prompts, raising urgent concerns about safety alignment in the event of misuse in the real world. Prior jailbreak attacks typically rewrite…

密码学与安全 · 计算机科学 2026-03-10 Moyang Chen , Zonghao Ying , Wenzhuo Xu , Quancheng Zou , Deyue Zhang , Dongdong Yang , Xiangzheng Zhang

Despite significant progress in semi-supervised learning for image object detection, several key issues are yet to be addressed for video object detection: (1) Achieving good performance for supervised video object detection greatly depends…

计算机视觉与模式识别 · 计算机科学 2023-11-14 Tanvir Mahmud , Chun-Hao Liu , Burhaneddin Yaman , Diana Marculescu

Temporal misalignment (time offset) between sensors is common in low cost visual-inertial odometry (VIO) systems. Such temporal misalignment introduces inconsistent constraints for state estimation, leading to a significant positioning…

机器人学 · 计算机科学 2024-03-20 Chaoran Xiong , Guoqing Liu , Qi Wu , Songpengcheng Xia , Tong Hua , Kehui Ma , Zhen Sun , Yan Xiang , Ling Pei

Multi-agents rely on accurate poses to share and align observations, enabling a collaborative perception of the environment. However, traditional GNSS-based localization often fails in GNSS-denied environments, making consistent feature…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Wenkai Lin , Qiming Xia , Wen Li , Xun Huang , Chenglu Wen

Applications in the Internet of Video Things (IoVT) domain have very tight constraints with respect to power and area. While neuromorphic vision sensors (NVS) may offer advantages over traditional imagers in this domain, the existing NVS…

图像与视频处理 · 电气工程与系统科学 2020-03-20 Deepak Singla , Soham Chatterjee , Lavanya Ramapantulu , Andres Ussa , Bharath Ramesh , Arindam Basu

This study presents a novel classroom surveillance system that integrates multiple modalities, including drowsiness, tracking of mobile phone usage, and face recognition,to assess student attentiveness with enhanced precision.The system…

计算机视觉与模式识别 · 计算机科学 2025-07-03 Ameer Hamza , Zuhaib Hussain But , Umar Arif , Samiya , M. Abdullah Asad , Muhammad Naeem

This study presents a comprehensive comparative analysis of two prominent self-supervised learning architectures for video action recognition: DINOv3, which processes frames independently through spatial feature extraction, and V-JEPA2,…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Sai Varun Kodathala , Rakesh Vunnam

The "You only look once v4"(YOLOv4) is one type of object detection methods in deep learning. YOLOv4-tiny is proposed based on YOLOv4 to simple the network structure and reduce parameters, which makes it be suitable for developing on the…

计算机视觉与模式识别 · 计算机科学 2022-07-21 Zicong Jiang , Liquan Zhao , Shuaiyang Li , Yanfei Jia

In recent years, surveillance cameras are widely deployed in public places, and the general crime rate has been reduced significantly due to these ubiquitous devices. Usually, these cameras provide cues and evidence after crimes are…

计算机视觉与模式识别 · 计算机科学 2020-10-20 Ming Cheng , Kunjing Cai , Ming Li

Predicting pedestrian movement is critical for human behavior analysis and also for safe and efficient human-agent interactions. However, despite significant advancements, it is still challenging for existing approaches to capture the…

计算机视觉与模式识别 · 计算机科学 2022-11-01 Pei Xu , Jean-Bernard Hayet , Ioannis Karamouzas

Deep learning models have enjoyed great success for image related computer vision tasks like image classification and object detection. For video related tasks like human action recognition, however, the advancements are not as significant…

计算机视觉与模式识别 · 计算机科学 2018-09-12 Xiaolin Song , Cuiling Lan , Wenjun Zeng , Junliang Xing , Jingyu Yang , Xiaoyan Sun

Object detection and segmentation are widely employed in computer vision applications, yet conventional models like YOLO series, while efficient and accurate, are limited by predefined categories, hindering adaptability in open scenarios.…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Ao Wang , Lihao Liu , Hui Chen , Zijia Lin , Jungong Han , Guiguang Ding

Reliable fall detection in elderly care requires monitoring systems that are not only accurate but also capable of producing stable, interpretable explanations of motion dynamics, a requirement that existing post hoc explainability methods…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Mohammad Saleh , Azadeh Tabatabaei

YOLOv8 plays a crucial role in the realm of autonomous driving, owing to its high-speed target detection, precise identification and positioning, and versatile compatibility across multiple platforms. By processing video streams or images…

计算机视觉与模式识别 · 计算机科学 2024-07-16 Zhipeng Ling , Qi Xin , Yiyu Lin , Guangze Su , Zuwei Shui

Speed bumps and potholes are the most common road anomalies, significantly affecting ride comfort and vehicle stability. Preview-based suspension control mitigates their impact by detecting such irregularities in advance and adjusting…

计算机视觉与模式识别 · 计算机科学 2025-10-08 Chuanqi Liang , Jie Fu , Miao Yu , Lei Luo

This research project aims to develop a real-time traffic sign detection system using the YOLOv5 architecture and deploy it for efficient traffic sign recognition during a drive in a suburban neighborhood. The project's primary objectives…

计算机视觉与模式识别 · 计算机科学 2023-10-17 Harish Loghashankar , Hieu Nguyen

Dynamic environments such as urban areas are still challenging for popular visual-inertial odometry (VIO) algorithms. Existing datasets typically fail to capture the dynamic nature of these environments, therefore making it difficult to…

机器人学 · 计算机科学 2021-02-12 Koji Minoda , Fabian Schilling , Valentin Wüest , Dario Floreano , Takehisa Yairi

Public spaces such as transport hubs, city centres, and event venues require timely and reliable detection of potentially violent behaviour to support public safety. While automated video analysis has made significant progress, practical…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Ganen Sethupathy , Lalit Dumka , Jan Schagen

This paper explores how deep learning techniques can improve visual-based SLAM performance in challenging environments. By combining deep feature extraction and deep matching methods, we introduce a versatile hybrid visual SLAM system…

机器人学 · 计算机科学 2024-06-05 Zhang Xiao , Shuaixin Li