中文
相关论文

相关论文: Vi-SAFE: A Spatial-Temporal Framework for Efficien…

200 篇论文

Recently, the domestic COVID-19 epidemic situation is serious, but in public places, some people do not wear masks or wear masks incorrectly, which requires the relevant staff to instantly remind and supervise them to wear masks correctly.…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Xuecheng Wu , Mengmeng Tian , Lanhang Zhai

Automatically detecting violence from surveillance footage is a subset of activity recognition that deserves special attention because of its wide applicability in unmanned security monitoring systems, internet video filtration, etc. In…

计算机视觉与模式识别 · 计算机科学 2021-10-08 Zahidul Islam , Mohammad Rukonuzzaman , Raiyan Ahmed , Md. Hasanul Kabir , Moshiur Farazi

Video Shadow Detection (VSD) aims to detect the shadow masks with frame sequence. Existing works suffer from inefficient temporal learning. Moreover, few works address the VSD problem by considering the characteristic (i.e., boundary) of…

计算机视觉与模式识别 · 计算机科学 2024-08-22 Haipeng Zhou , Honqiu Wang , Tian Ye , Zhaohu Xing , Jun Ma , Ping Li , Qiong Wang , Lei Zhu

Lidar-based SLAM systems are highly sensitive to adverse conditions such as occlusion, noise, and field-of-view (FoV) degradation, yet existing robustness evaluation methods either lack physical grounding or do not capture sensor-specific…

机器人学 · 计算机科学 2025-12-10 Doumegna Mawuto Koudjo Felix , Xianjia Yu , Zhuo Zou , Tomi Westerlund

Urban traffic environments present unique challenges for object detection, particularly with the increasing presence of micromobility vehicles like e-scooters and bikes. To address this object detection problem, this work introduces an…

计算机视觉与模式识别 · 计算机科学 2024-10-31 Khalil Sabri , Célia Djilali , Guillaume-Alexandre Bilodeau , Nicolas Saunier , Wassim Bouachir

Time-Sensitive Networking (TSN) is an emerging real-time Ethernet technology that provides deterministic communication for time-critical traffic. At its core, TSN relies on Time-Aware Shaper (TAS) for pre-allocating frames in specific time…

网络与互联网体系结构 · 计算机科学 2024-03-05 Xuyan Jiang , Xiangrui Yang , Tongqing Zhou , Wenwen Fu , Wei Quan , Yihao Jiao , Yinhan Sun , Zhigang Sun

In recent years, event cameras (DVS - Dynamic Vision Sensors) have been used in vision systems as an alternative or supplement to traditional cameras. They are characterised by high dynamic range, high temporal resolution, low latency, and…

计算机视觉与模式识别 · 计算机科学 2022-11-28 Piotr Wzorek , Tomasz Kryjak

This paper presents a new method for anomaly detection in automated systems with time and compute sensitive requirements, such as autonomous driving, with unparalleled efficiency. As systems like autonomous driving become increasingly…

计算机视觉与模式识别 · 计算机科学 2025-03-12 Andrew Gao , Jun Liu

Visual place recognition (VPR) using deep networks has achieved state-of-the-art performance. However, most of them require a training set with ground truth sensor poses to obtain positive and negative samples of each observation's spatial…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Chao Chen , Zegang Cheng , Xinhao Liu , Yiming Li , Li Ding , Ruoyu Wang , Chen Feng

Video foreground segmentation (VFS) is an important computer vision task wherein one aims to segment the objects under motion from the background. Most of the current methods are image-based, i.e., rely only on spatial cues while ignoring…

计算机视觉与模式识别 · 计算机科学 2024-02-05 Praveen Kumar Pokala , Jaya Sai Kiran Patibandla , Naveen Kumar Pandey , Balakrishna Reddy Pailla

Recently, video object segmentation (VOS) networks typically use memory-based methods: for each query frame, the mask is predicted by space-time matching to memory frames. Despite these methods having superior performance, they suffer from…

计算机视觉与模式识别 · 计算机科学 2024-05-08 Yadang Chen , Wentao Zhu , Zhi-Xin Yang , Enhua Wu

Traffic light and sign recognition are key for Autonomous Vehicles (AVs) because perception mistakes directly influence navigation and safety. In addition to digital adversarial attacks, models are vulnerable to existing perturbations…

计算机视觉与模式识别 · 计算机科学 2025-10-06 Abhishek Joshi , Jahnavi Krishna Koda , Abhishek Phadke

Vision-Language Models (VLMs) exhibit impressive performance, yet the integration of powerful vision encoders has significantly broadened their attack surface, rendering them increasingly susceptible to jailbreak attacks. However, lacking…

计算机视觉与模式识别 · 计算机科学 2026-02-26 Jiaxin Song , Yixu Wang , Jie Li , Rui Yu , Yan Teng , Xingjun Ma , Yingchun Wang

Although large vision-language-action (VLA) models pretrained on extensive robot datasets offer promising generalist policies for robotic learning, they still struggle with spatial-temporal dynamics in interactive robotics, making them less…

机器人学 · 计算机科学 2025-06-09 Ruijie Zheng , Yongyuan Liang , Shuaiyi Huang , Jianfeng Gao , Hal Daumé , Andrey Kolobov , Furong Huang , Jianwei Yang

Visual object detection utilizing deep learning plays a vital role in computer vision and has extensive applications in transportation engineering. This paper focuses on detecting pavement marking quality during daytime using the You Only…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Gian Antariksa , Rohit Chakraborty , Shriyank Somvanshi , Subasish Das , Mohammad Jalayer , Deep Rameshkumar Patel , David Mills

Detecting and recognizing human action in videos with crowded scenes is a challenging problem due to the complex environment and diversity events. Prior works always fail to deal with this problem in two aspects: (1) lacking utilizing…

计算机视觉与模式识别 · 计算机科学 2020-10-19 Li Yuan , Yichen Zhou , Shuning Chang , Ziyuan Huang , Yunpeng Chen , Xuecheng Nie , Tao Wang , Jiashi Feng , Shuicheng Yan

The YOLO (You Only Look Once) series has been a leading framework in real-time object detection, consistently improving the balance between speed and accuracy. However, integrating attention mechanisms into YOLO has been challenging due to…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Rahima Khanam , Muhammad Hussain

Unsupervised Video Object Segmentation (VOS) aims at identifying the contours of primary foreground objects in videos without any prior knowledge. However, previous methods do not fully use spatial-temporal context and fail to tackle this…

计算机视觉与模式识别 · 计算机科学 2023-11-21 Ping Li , Yu Zhang , Li Yuan , Huaxin Xiao , Binbin Lin , Xianghua Xu

Real-time hand tracking in trauma surgery is essential for supporting rapid and precise intraoperative decisions. We propose a YOLOv10-based framework that simultaneously localizes hands and classifies their laterality (left or right) in…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Kedi Sun , Le Zhang

Spatio-temporal prediction is a crucial research area in data-driven urban computing, with implications for transportation, public safety, and environmental monitoring. However, scalability and generalization challenges remain significant…

机器学习 · 计算机科学 2024-09-12 Jiabin Tang , Wei Wei , Lianghao Xia , Chao Huang