中文
相关论文

相关论文: Video Labeling for Automatic Video Surveillance in…

200 篇论文

Unmanned Aerial Vehicles (UAVs) offer wide-ranging applications but also pose significant safety and privacy violation risks in areas like airport and infrastructure inspection, spurring the rapid development of Anti-UAV technologies in…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Chunhui Zhang , Li Liu , Zhipeng Zhang , Yong Wang , Hao Wen , Xi Zhou , Shiming Ge , Yanfeng Wang

With the rapid advancement of UAV technology and its extensive application in various fields such as military reconnaissance, environmental monitoring, and logistics, achieving efficient and accurate Anti-UAV tracking has become essential.…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Guanghai Ding , Yihua Ren , Yuting Liu , Qijun Zhao , Shuiwang Li

We introduce Latent Action Pretraining for general Action models (LAPA), an unsupervised method for pretraining Vision-Language-Action (VLA) models without ground-truth robot action labels. Existing Vision-Language-Action models require…

Precise vehicle state estimation is crucial for safe and reliable autonomous driving. The number of measurable states and their precision offered by the onboard vehicle sensor system are often constrained by cost. For instance, measuring…

A generalist robot should perform effectively across various environments. However, most existing approaches heavily rely on scaling action-annotated data to enhance their capabilities. Consequently, they are often limited to single…

机器人学 · 计算机科学 2025-11-04 Qingwen Bu , Yanting Yang , Jisong Cai , Shenyuan Gao , Guanghui Ren , Maoqing Yao , Ping Luo , Hongyang Li

Filming sport videos from an aerial view has always been a hard and an expensive task to achieve, especially in sports that require a wide open area for its normal development or the ones that put in danger human safety. Recently, a new…

机器人学 · 计算机科学 2022-12-23 Dennis Casazola , Fabio Arnez , Huascar Espinoza

VERSA provides a general-purpose framework for defining and recognizing events in live or recorded surveillance video streams. The approach for event recognition in VERSA is using a declarative logic language to define the spatial and…

计算机视觉与模式识别 · 计算机科学 2010-07-23 Stephen O'Hara

Vision-language-action models (VLAs) have garnered significant attention for their potential in advancing robotic manipulation. However, previous approaches predominantly rely on the general comprehension capabilities of vision-language…

计算机视觉与模式识别 · 计算机科学 2025-06-25 Yuqi Wang , Xinghang Li , Wenxuan Wang , Junbo Zhang , Yingyan Li , Yuntao Chen , Xinlong Wang , Zhaoxiang Zhang

Smart traffic engineering and intelligent transportation services are in increasing demand from governmental authorities to optimize traffic performance and thus reduce energy costs, increase the drivers' safety and comfort, ensure traffic…

计算机视觉与模式识别 · 计算机科学 2023-03-02 Bilel Benjdira , Anis Koubaa , Ahmad Taher Azar , Zahid Khan , Adel Ammar , Wadii Boulila

Large-scale semantic image annotation is a significant challenge for learning-based perception systems in robotics. Current approaches often rely on human labelers, which can be expensive, or simulation data, which can visually or…

计算机视觉与模式识别 · 计算机科学 2022-03-15 Brijen Thananjeyan , Justin Kerr , Huang Huang , Joseph E. Gonzalez , Ken Goldberg

Video anomaly detection (VAD) in autonomous driving scenario is an important task, however it involves several challenges due to the ego-centric views and moving camera. Due to this, it remains largely under-explored. While recent…

计算机视觉与模式识别 · 计算机科学 2024-08-13 Utkarsh Tiwari , Snehashis Majhi , Michal Balazia , François Brémond

The growing demand for surveillance in public spaces presents significant challenges due to the shortage of human resources. Current AI-based video surveillance systems heavily rely on core computer vision models that require extensive…

计算机视觉与模式识别 · 计算机科学 2025-03-19 Joao Pereira , Vasco Lopes , David Semedo , Joao Neves

Along with the increasing use of unmanned aerial vehicles (UAVs), large volumes of aerial videos have been produced. It is unrealistic for humans to screen such big data and understand their contents. Hence methodological research on the…

计算机视觉与模式识别 · 计算机科学 2020-06-26 Lichao Mou , Yuansheng Hua , Pu Jin , Xiao Xiang Zhu

Live tracking of wildlife via high-resolution video processing directly onboard drones is widely unexplored and most existing solutions rely on streaming video to ground stations to support navigation. Yet, both autonomous animal-reactive…

计算机视觉与模式识别 · 计算机科学 2025-05-26 Nguyen Ngoc Dat , Tom Richardson , Matthew Watson , Kilian Meier , Jenna Kline , Sid Reid , Guy Maalouf , Duncan Hine , Majid Mirmehdi , Tilo Burghardt

Deep learning algorithms have pushed the boundaries of computer vision research and have depicted commendable performance in a variety of applications. However, training a robust deep neural network necessitates a large amount of labeled…

计算机视觉与模式识别 · 计算机科学 2023-07-13 Debanjan Goswami , Shayok Chakraborty

Recent years have witnessed the rapid progress of perception algorithms on top of LiDAR, a widely adopted sensor for autonomous driving systems. These LiDAR-based solutions are typically data hungry, requiring a large amount of data to be…

机器人学 · 计算机科学 2020-11-23 Tai Wang , Conghui He , Zhe Wang , Jianping Shi , Dahua Lin

The proliferation of unmanned aerial vehicles (UAVs) in controlled airspace presents significant risks, including potential collisions, disruptions to air traffic, and security threats. Ensuring the safe and efficient operation of airspace,…

机器人学 · 计算机科学 2025-10-22 Francisco Giral , Ignacio Gómez , Soledad Le Clainche

Internet of Things (IoT) involves sensors for monitoring and wireless networks for efficient communication. However, resource-constrained IoT devices and limitations in existing wireless technologies hinder its full potential. Integrating…

网络与互联网体系结构 · 计算机科学 2023-11-10 Poorvi Joshi , Alakesh Kalita , Mohan Gurusamy

Object detection from images captured by Unmanned Aerial Vehicles (UAVs) is becoming increasingly useful. Despite the great success of the generic object detection methods trained on ground-to-ground images, a huge performance drop is…

计算机视觉与模式识别 · 计算机科学 2020-10-06 Zhenyu Wu , Karthik Suresh , Priya Narayanan , Hongyu Xu , Heesung Kwon , Zhangyang Wang

Large-scale Vision Language Models (LVLMs) exhibit advanced capabilities in tasks that require visual information, including object detection. These capabilities have promising applications in various industrial domains, such as autonomous…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Haruki Sakajo , Hiroshi Takato , Hiroshi Tsutsui , Komei Soda , Hidetaka Kamigaito , Taro Watanabe