中文
相关论文

相关论文: TrajSV: A Trajectory-based Model for Sports Video …

200 篇论文

A new approach in team sports analysis consists in studying positioning and movements of players during the game in relation to team performance. State of the art tracking systems produce spatio-temporal traces of players that have…

应用统计 · 统计学 2017-07-06 Rodolfo Metulini , Marica Manisera , Paola Zuccolotto

To date, machine learning for human action recognition in video has been widely implemented in sports activities. Although some studies have been successful in the past, precision is still the most significant concern. In this study, we…

计算机视觉与模式识别 · 计算机科学 2020-12-02 Cheng Yan , Xin Li , Guoqiang Li

Learning manipulation skills from human demonstration videos presents a promising yet challenging problem, primarily due to the significant embodiment gap between human body and robot manipulators. Existing methods rely on paired datasets…

机器人学 · 计算机科学 2025-10-10 YuHang Tang , Yixuan Lou , Pengfei Han , Haoming Song , Xinyi Ye , Dong Wang , Bin Zhao

The prosperity of Multimodal Large Language Models (MLLMs) has stimulated the demand for video reasoning segmentation, which aims to segment video objects based on human instructions. Previous studies rely on unidirectional and implicit…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Jingnan Luo , Mingqi Gao , Jun Liu , Bin-Bin Gao , Feng Zheng

Masked video modeling (MVM) has emerged as a simple and scalable self-supervised pretraining paradigm, but only encodes motion information implicitly, limiting the encoding of temporal dynamics in the learned representations. As a result,…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Renaud Vandeghen , Fida Mohammad Thoker , Marc Van Droogenbroeck , Bernard Ghanem

Pedestrian trajectory prediction is crucial for several applications such as robotics and self-driving vehicles. Significant progress has been made in the past decade thanks to the availability of pedestrian trajectory datasets, which…

计算机视觉与模式识别 · 计算机科学 2024-11-04 Pranav Singh Chib , Pravendra Singh

Semantic representation is of great benefit to the video text tracking(VTT) task that requires simultaneously classifying, detecting, and tracking texts in the video. Most existing approaches tackle this task by appearance similarity in…

计算机视觉与模式识别 · 计算机科学 2022-08-22 Zhuang Li , Weijia Wu , Mike Zheng Shou , Jiahong Li , Size Li , Zhongyuan Wang , Hong Zhou

Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limited. In this paper, we first introduce **TripVVT-10K**, the largest and most diverse…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Dingbao Shao , Song Wu , Shenyi Wang , Ye Wang , Ziheng Tang , Fei Liu , Jiang Lin , Xinyu Chen , Qian Wang , Ying Tai , Jian Yang , Zili Yi

Discriminative representation is crucial for the association step in multi-object tracking. Recent work mainly utilizes features in single or neighboring frames for constructing metric loss and empowering networks to extract representation…

计算机视觉与模式识别 · 计算机科学 2022-04-06 En Yu , Zhuoling Li , Shoudong Han

Multi-object tracking (MOT) is a critical and challenging task in computer vision, particularly in situations involving objects with similar appearances but diverse movements, as seen in team sports. Current methods, largely reliant on…

计算机视觉与模式识别 · 计算机科学 2024-04-23 Atom Scott , Ikuma Uchida , Ning Ding , Rikuhei Umemoto , Rory Bunker , Ren Kobayashi , Takeshi Koyama , Masaki Onishi , Yoshinari Kameda , Keisuke Fujii

Current video generation models produce physically inconsistent motion that violates real-world dynamics. We propose TrajVLM-Gen, a two-stage framework for physics-aware image-to-video generation. First, we employ a Vision Language Model to…

计算机视觉与模式识别 · 计算机科学 2025-10-02 Fan Yang , Zhiyang Chen , Yousong Zhu , Xin Li , Jinqiao Wang

Robust ball tracking under occlusion remains a key challenge in sports video analysis, affecting tasks like event detection and officiating. We present TOTNet, a Temporal Occlusion Tracking Network that leverages 3D convolutions,…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Hao Xu , Arbind Agrahari Baniya , Sam Wells , Mohamed Reda Bouadjenek , Richard Dazely , Sunil Aryal

Sports action classification representing complex body postures and player-object interactions is an emerging area in image-based sports analysis. Some works have contributed to automated sports action recognition using machine learning…

计算机视觉与模式识别 · 计算机科学 2025-07-16 Palash Ray , Mahuya Sasmal , Asish Bera

Soccer analytics is attracting increasing interest in academia and industry, thanks to the availability of data that describe all the spatio-temporal events that occur in each match. These events (e.g., passes, shots, fouls) are collected…

计算机视觉与模式识别 · 计算机科学 2020-07-14 Danilo Sorano , Fabio Carrara , Paolo Cintia , Fabrizio Falchi , Luca Pappalardo

Visual saliency prediction using transformers - Convolutional neural networks (CNNs) have significantly advanced computational modelling for saliency prediction. However, accurately simulating the mechanisms of visual attention in the human…

多媒体 · 计算机科学 2022-06-30 Jianxun Lou , Hanhe Lin , David Marshall , Dietmar Saupe , Hantao Liu

To understand human behaviors, action recognition based on videos is a common approach. Compared with image-based action recognition, videos provide much more information. Reducing the ambiguity of actions and in the last decade, many works…

计算机视觉与模式识别 · 计算机科学 2022-06-03 Fei Wu , Qingzhong Wang , Jian Bian , Haoyi Xiong , Ning Ding , Feixiang Lu , Jun Cheng , Dejing Dou

Due to the advent of new mobile devices and tracking sensors in recent years, huge amounts of data are being produced every day. Therefore, novel methodologies need to emerge that dive through this vast sea of information and generate…

计算机视觉与模式识别 · 计算机科学 2022-05-31 Ioannis Kontopoulos , Antonios Makris , Konstantinos Tserpes , Vania Bogorny

Demystifying the interactions among multiple agents from their past trajectories is fundamental to precise and interpretable trajectory prediction. However, previous works only consider pair-wise interactions with limited relational…

计算机视觉与模式识别 · 计算机科学 2022-04-21 Chenxin Xu , Maosen Li , Zhenyang Ni , Ya Zhang , Siheng Chen

Current approaches for video grounding propose kinds of complex architectures to capture the video-text relations, and have achieved impressive improvements. However, it is hard to learn the complicated multi-modal relations by only…

计算机视觉与模式识别 · 计算机科学 2021-08-25 Xinpeng Ding , Nannan Wang , Shiwei Zhang , De Cheng , Xiaomeng Li , Ziyuan Huang , Mingqian Tang , Xinbo Gao

Human action recognition in video is an active yet challenging research topic due to high variation and complexity of data. In this paper, a novel video based action recognition framework utilizing complementary cues is proposed to handle…

计算机视觉与模式识别 · 计算机科学 2019-09-10 Muhammad Usman Khalid , Jie Yu