中文
相关论文

相关论文: Multi-Modal Soccer Scene Analysis with Masked Pre-…

200 篇论文

The recently proposed action spotting task consists in finding the exact timestamp in which an event occurs. This task fits particularly well for soccer videos, where events correspond to salient actions strictly defined by soccer rules (a…

计算机视觉与模式识别 · 计算机科学 2021-02-16 Matteo Tomei , Lorenzo Baraldi , Simone Calderara , Simone Bronzin , Rita Cucchiara

In this paper we propose a system capable of tracking multiple soccer players in different types of video quality. The main goal, in contrast to most state-of-art soccer player tracking systems, is the ability of execute effectively…

计算机视觉与模式识别 · 计算机科学 2021-05-25 Eloi Martins , José Henrique Brito

Toward the goal of automatic production for sports broadcasts, a paramount task consists in understanding the high-level semantic information of the game in play. For instance, recognizing and localizing the main actions of the game would…

计算机视觉与模式识别 · 计算机科学 2021-04-15 Silvio Giancola , Bernard Ghanem

This paper presents CourtMotion, a spatiotemporal modeling framework for analyzing and predicting game events and plays as they develop in professional basketball. Anticipating basketball events requires understanding both physical motion…

计算机视觉与模式识别 · 计算机科学 2025-12-10 Omer Sela , Michael Chertok , Lior Wolf

We modeled the dynamics of a soccer match based on a network representation where players are nodes discretely clustered into homogeneous groups. Players were grouped by physical proximity, supported by the intuitive notion that competing…

社会与信息网络 · 计算机科学 2021-08-09 Luis Ramada Pereira , Rui J. Lopes , Jorge Louçã , Duarte Araújo , João Ramos

We present the Transparent Earth, a transformer-based architecture for reconstructing subsurface properties from heterogeneous datasets that vary in sparsity, resolution, and modality, where each modality represents a distinct type of…

机器学习 · 计算机科学 2025-09-24 Arnab Mazumder , Javier E. Santos , Noah Hobbs , Mohamed Mehana , Daniel O'Malley

SoccerTrack v2 is a new public dataset for advancing multi-object tracking (MOT), game state reconstruction (GSR), and ball action spotting (BAS) in soccer analytics. Unlike prior datasets that use broadcast views or limited scenarios,…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Atom Scott , Ikuma Uchida , Kento Kuroda , Yufi Kim , Keisuke Fujii

Multi-modal 3D object understanding has gained significant attention, yet current approaches often assume complete data availability and rigid alignment across all modalities. We present CrossOver, a novel framework for cross-modal 3D scene…

计算机视觉与模式识别 · 计算机科学 2025-04-08 Sayan Deb Sarkar , Ondrej Miksik , Marc Pollefeys , Daniel Barath , Iro Armeni

The integration of artificial intelligence in sports analytics has transformed soccer video understanding, enabling real-time, automated insights into complex game dynamics. Traditional approaches rely on isolated data streams, limiting…

计算机视觉与模式识别 · 计算机科学 2025-05-23 Sushant Gautam , Cise Midoglu , Vajira Thambawita , Michael A. Riegler , Pål Halvorsen , Mubarak Shah

Multi-modal 3D scene understanding has gained considerable attention due to its wide applications in many areas, such as autonomous driving and human-computer interaction. Compared to conventional single-modal 3D understanding, introducing…

计算机视觉与模式识别 · 计算机科学 2023-10-25 Yinjie Lei , Zixuan Wang , Feng Chen , Guoqing Wang , Peng Wang , Yang Yang

Context plays a significant role in the generation of motion for dynamic agents in interactive environments. This work proposes a modular method that utilises a learned model of the environment for motion prediction. This modularity…

机器学习 · 计算机科学 2021-01-05 Todor Davchev , Michael Burke , Subramanian Ramamoorthy

Camera calibration and localization, sometimes simply named camera calibration, enables many applications in the context of soccer broadcasting, for instance regarding the interpretation and analysis of the game, or the insertion of…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Floriane Magera , Thomas Hoyoux , Olivier Barnich , Marc Van Droogenbroeck

Intelligent sports video analysis demands a comprehensive understanding of temporal context, from micro-level actions to macro-level game strategies. Existing end-to-end models often struggle with this temporal hierarchy, offering solutions…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Tsz-To Wong , Ching-Chun Huang , Hong-Han Shuai

Anticipating future actions is a highly challenging task due to the diversity and scale of potential future actions; yet, information from different modalities help narrow down plausible action choices. Each modality can provide diverse and…

计算机视觉与模式识别 · 计算机科学 2024-08-30 Apoorva Beedu , Harish Haresamudram , Karan Samel , Irfan Essa

In a soccer game, the information provided by detecting and tracking brings crucial clues to further analyze and understand some tactical aspects of the game, including individual and team actions. State-of-the-art tracking algorithms…

计算机视觉与模式识别 · 计算机科学 2020-11-23 Samuel Hurault , Coloma Ballester , Gloria Haro

Team sports represent complex phenomena characterized by both spatial and temporal dimensions, making their analysis inherently challenging. In this study, we examine team sports as complex systems, specifically focusing on the tactical…

Significant development of communication technology over the past few years has motivated research in multi-modal summarization techniques. A majority of the previous works on multi-modal summarization focus on text and images. In this…

信息检索 · 计算机科学 2020-05-20 Anubhav Jangra , Sriparna Saha , Adam Jatowt , Mohammad Hasanuzzaman

Clubs with access to expensive multi-camera setups or GPS tracking systems gain a competitive advantage through detailed data, whereas lower-budget teams are often unable to collect similar information. This paper examines whether such data…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Daniel Tshiani

This study developed a new explainable artificial intelligence algorithm called PassAI, which classifies successful or failed passes in a soccer game and explains its rationale using both tracking and passer's seasonal stats information.…

人机交互 · 计算机科学 2026-04-06 Ryota Takamido , Jun Ota , Hiroki Nakamoto

Attention-based models are appealing for multimodal processing because inputs from multiple modalities can be concatenated and fed to a single backbone network - thus requiring very little fusion engineering. The resulting representations…