中文
相关论文

相关论文: SoccerNet 2022 Challenges Results

200 篇论文

Recently, video text detection, tracking, and recognition in natural scenes are becoming very popular in the computer vision community. However, most existing algorithms and benchmarks focus on common text cases (e.g., normal size, density)…

计算机视觉与模式识别 · 计算机科学 2023-04-11 Weijia Wu , Yuzhong Zhao , Zhuang Li , Jiahong Li , Mike Zheng Shou , Umapada Pal , Dimosthenis Karatzas , Xiang Bai

One of the requirements for team sports analysis is to track and recognize players. Many tracking and reidentification methods have been proposed in the context of video surveillance. They show very convincing results when tested on public…

计算机视觉与模式识别 · 计算机科学 2022-04-11 Adrien Maglo , Astrid Orcesi , Quoc-Cuong Pham

Multi-Object Tracking over humans has improved rapidly with the development of object detection and re-identification. However, multi-actor tracking over humans with similar appearance and nonlinear movement can still be very challenging…

计算机视觉与模式识别 · 计算机科学 2022-09-28 Hsiang-Wei Huang , Cheng-Yen Yang , Jenq-Neng Hwang , Pyong-Kun Kim , Kwangju Kim , Kyoungoh Lee

We describe SemEval-2022 Task 7, a shared task on rating the plausibility of clarifications in instructional texts. The dataset for this task consists of manually clarified how-to guides for which we generated alternative clarifications and…

计算与语言 · 计算机科学 2023-09-22 Michael Roth , Talita Anthonio , Anna Sauer

Automated soccer commentary generation has evolved from template-based systems to advanced neural architectures, aiming to produce real-time descriptions of sports events. While frameworks like SoccerNet-Caption laid foundational work,…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Chidaksh Ravuru

This technical report presents a brief description of our submission to the dense video captioning task of ActivityNet Challenge 2020. Our approach follows a two-stage pipeline: first, we extract a set of temporal event proposals; then we…

计算机视觉与模式识别 · 计算机科学 2020-08-13 Teng Wang , Huicheng Zheng , Mingjing Yu

How can we effectively engineer a computer vision system that is able to interpret videos from unconstrained mobility platforms like UAVs? One promising option is to make use of image restoration and enhancement algorithms from the area of…

计算机视觉与模式识别 · 计算机科学 2020-11-23 Sreya Banerjee , Rosaura G. VidalMata , Zhangyang Wang , Walter J. Scheirer

This paper explores the real-time summarization of scheduled events such as soccer games from torrential flows of Twitter streams. We propose and evaluate an approach that substantially shrinks the stream of tweets in real-time, and…

信息检索 · 计算机科学 2012-04-18 Arkaitz Zubiaga , Damiano Spina , Enrique Amigó , Julio Gonzalo

This paper presents the baseline method proposed for the Sports Video task part of the MediaEval 2021 benchmark. This task proposes a stroke detection and a stroke classification subtasks. This baseline addresses both subtasks. The…

计算机视觉与模式识别 · 计算机科学 2021-12-23 Pierre-Etienne Martin

The Video Assistant Referee (VAR) has revolutionized association football, enabling referees to review incidents on the pitch, make informed decisions, and ensure fairness. However, due to the lack of referees in many countries and the high…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Jan Held , Anthony Cioppa , Silvio Giancola , Abdullah Hamdi , Bernard Ghanem , Marc Van Droogenbroeck

Video object segmentation can be considered as one of the most challenging computer vision problems. Indeed, so far, no existing solution is able to effectively deal with the peculiarities of real-world videos, especially in cases of…

计算机视觉与模式识别 · 计算机科学 2016-01-06 Simone Palazzo , Concetto Spampinato , Daniela Giordano

Multimedia event detection is the task of detecting a specific event of interest in an user-generated video on websites. The most fundamental challenge facing this task lies in the enormously varying quality of the video as well as the…

计算机视觉与模式识别 · 计算机科学 2021-10-18 Minnan Luo , Xiaojun Chang , Chen Gong

While existing video benchmarks largely consider specialized downstream tasks like retrieval or question-answering (QA), contemporary multimodal AI systems must be capable of well-rounded common-sense reasoning akin to human visual…

计算机视觉与模式识别 · 计算机科学 2024-06-17 Kate Sanders , Benjamin Van Durme

This paper introduces the methods and the results of AIM 2022 challenge on Instagram Filter Removal. Social media filters transform the images by consecutive non-linear operations, and the feature maps of the original content may be…

In this work we propose a multi-modal architecture for analyzing soccer scenes from tactical camera footage, with a focus on three core tasks: ball trajectory inference, ball state classification, and ball possessor identification. To this…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Marc Peral , Guillem Capellera , Luis Ferraz , Antonio Rubio , Antonio Agudo

Soccer Simulation 2D (SS2D) is a simulation of a real soccer game in two dimensions. In soccer, passing behavior is an essential action for keeping the ball in possession of our team and creating goal opportunities. Similarly, for SS2D,…

人工智能 · 计算机科学 2024-01-09 Nader Zare , Mahtab Sarvmaili , Aref Sayareh , Omid Amini , Stan Matwin Amilcar Soares

The proliferation of generative video technologies has intensified the need for reliable methods to detect and characterize synthetic media. To address this challenge, we organized the \href{https://safe-video-2025.dsri.org}{SAFE: Synthetic…

Traditional approaches to measuring visual exploratory behavior in soccer rely on counting visual exploratory actions (VEAs) based on rapid head movements exceeding 125{\deg}/s, but this method suffer from player position bias (i.e., a…

机器学习 · 计算机科学 2026-02-24 Joris Bekkers

This paper presents the concepts of Artificial Intelligence, Multi-Agent-Systems, Coordination, Intelligent Robotics and Deep Reinforcement Learning. Emphasis is given on and how AI and DRL, may be efficiently used to create efficient robot…

机器人学 · 计算机科学 2023-12-29 Luis Paulo Reis

While recent video world models can generate highly realistic videos, their ability to perform semantic reasoning and planning remains unclear and unquantified. We introduce Target-Bench, the first benchmark that enables comprehensive…