中文
相关论文

相关论文: Automatic Observer Script for StarCraft: Brood War…

200 篇论文

Automated live visual descriptions can aid blind people in understanding their surroundings with autonomy and independence. However, providing descriptions that are rich, contextual, and just-in-time has been a long-standing challenge in…

人机交互 · 计算机科学 2024-08-14 Ruei-Che Chang , Yuxuan Liu , Anhong Guo

We present a method for fast training of vision based control policies on real robots. The key idea behind our method is to perform multi-task Reinforcement Learning with auxiliary tasks that differ not only in the reward to be optimized…

We describe the development of a system for an automated, iterative, real-time classification of transient events discovered in synoptic sky surveys. The system under development incorporates a number of Machine Learning techniques, mostly…

天体物理仪器与方法 · 物理学 2011-10-24 S. G. Djorgovski , C. Donalek , A. Mahabal , B. Moghaddam , M. Turmon , M. Graham , A. Drake , N. Sharma , Y. Chen

This paper presents a semi-supervised approach to extracting and analyzing combat phases in judo tournaments using live-streamed footage. The objective is to automate the annotation and summarization of live streamed judo matches. We train…

计算机视觉与模式识别 · 计算机科学 2024-12-11 Anthony Miyaguchi , Jed Moutahir , Tanmay Sutar

Screen recordings of mobile applications are easy to capture and include a wealth of information, making them a popular mechanism for users to inform developers of the problems encountered in the bug reports. However, watching the bug…

软件工程 · 计算机科学 2023-02-03 Sidong Feng , Mulong Xie , Yinxing Xue , Chunyang Chen

Video captioning aims to automatically generate natural language descriptions of video content, which has drawn a lot of attention recent years. Generating accurate and fine-grained captions needs to not only understand the global content…

计算机视觉与模式识别 · 计算机科学 2019-06-12 Junchao Zhang , Yuxin Peng

Concrete workability is essential for construction quality, with the slump test being the most widely used on-site method for its assessment. However, traditional slump testing is manual, time-consuming, and highly operator-dependent,…

计算机视觉与模式识别 · 计算机科学 2026-01-15 Youngmin Kim , Giyeong Oh , Kwangsoo Youm , Youngjae Yu

To meet the growing demand for systematic surgical training, wet-lab environments have become indispensable platforms for hands-on practice in ophthalmology. Yet, traditional wet-lab training depends heavily on manual performance…

计算机视觉与模式识别 · 计算机科学 2025-10-07 Negin Ghamsarian , Raphael Sznitman , Klaus Schoeffmann , Jens Kowal

In this paper, we present our approach for the Track 1 of the Chinese Auditory Attention Decoding (Chinese AAD) Challenge at ISCSLP 2024. Most existing spatial auditory attention decoding (Sp-AAD) methods employ an isolated window…

声音 · 计算机科学 2024-08-28 Zelin Qiu , Dingding Yao , Junfeng Li

Autonomous vehicles navigate in dynamically changing environments under a wide variety of conditions, being continuously influenced by surrounding objects. Modelling interactions among agents is essential for accurately forecasting other…

机器学习 · 计算机科学 2021-06-01 Sandra Carrasco , David Fernández Llorca , Miguel Ángel Sotelo

We study the fundamental problem of butterfly (i.e. (2,2)-bicliques) counting in bipartite streaming graphs. Similar to triangles in unipartite graphs, enumerating butterflies is crucial in understanding the structure of bipartite graphs.…

数据库 · 计算机科学 2021-02-04 Aida Sheshbolouki , M. Tamer Özsu

Motivated by vision-based reinforcement learning (RL) problems, in particular Atari games from the recent benchmark Aracade Learning Environment (ALE), we consider spatio-temporal prediction problems where future (image-)frames are…

机器学习 · 计算机科学 2015-12-23 Junhyuk Oh , Xiaoxiao Guo , Honglak Lee , Richard Lewis , Satinder Singh

Visual monitoring of industrial assembly tasks is critical for preventing equipment damage due to procedural errors and ensuring worker safety. Although commercial solutions exist, they typically require rigid workspace setups or the…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Mattia Nardon , Stefano Messelodi , Antonio Granata , Fabio Poiesi , Alberto Danese , Davide Boscaini

Visual attention serves as a means of feature selection mechanism in the perceptual system. Motivated by Broadbent's leaky filter model of selective attention, we evaluate how such mechanism could be implemented and affect the learning…

机器学习 · 计算机科学 2020-06-19 Liu Yuezhang , Ruohan Zhang , Dana H. Ballard

We consider a multi-adversary version of the supervisory control problem for discrete-event systems, in which an adversary corrupts the observations available to the supervisor. The supervisor's goal is to enforce a specific language in…

系统与控制 · 计算机科学 2018-08-23 Masashi Wakaiki , Paulo Tabuada , Joao P. Hespanha

In the process of making a movie, directors constantly care about where the spectator will look on the screen. Shot composition, framing, camera movements or editing are tools commonly used to direct attention. In order to provide a…

计算机视觉与模式识别 · 计算机科学 2021-03-01 Alexandre Bruckert , Marc Christie , Olivier Le Meur

In this paper, we focus on the problem of applying the transformer structure to video captioning effectively. The vanilla transformer is proposed for uni-modal language generation task such as machine translation. However, video captioning…

计算机视觉与模式识别 · 计算机科学 2020-07-24 Tao Jin , Siyu Huang , Ming Chen , Yingming Li , Zhongfei Zhang

Artificial visual attention systems aim to support technical systems in visual tasks by applying the concepts of selective attention observed in humans and other animals. Such systems are typically evaluated against ground truth obtained…

计算机视觉与模式识别 · 计算机科学 2013-08-01 Jan Tünnermann , Markus Hennig , Michael Silbernagel , Bärbel Mertsching

Automatically detecting and classifying strokes in table tennis video can streamline training workflows, enrich broadcast overlays, and enable fine-grained performance analytics. For this to be possible, annotated video data of table tennis…

计算机视觉与模式识别 · 计算机科学 2026-01-09 Moamal Fadhil Abdul-Mahdi , Jonas Bruun Hubrechts , Thomas Martini Jørgensen , Emil Hovad

Click-point-based interactive segmentation has received widespread attention due to its efficiency. However, it's hard for existing algorithms to obtain precise and robust responses after multiple clicks. In this case, the segmentation…

计算机视觉与模式识别 · 计算机科学 2024-05-08 Long Xu , Yongquan Chen , Rui Huang , Feng Wu , Shiwu Lai