中文
相关论文

相关论文: Exploring an Infinite Space with Finite Memory Sco…

200 篇论文

The first passage search of a diffusing target (prey) by multiple searchers (predators) in confinement is an important problem in the stochastic process literature. While the analogous problem in open space has been studied in some details,…

统计力学 · 物理学 2021-01-04 Indrani Nayak , Amitabha Nandi , Dibyendu Das

We consider two-player games over graphs and give tight bounds on the memory size of strategies ensuring safety objectives. More specifically, we show that the minimal number of memory states of a strategy ensuring a safety objective is…

计算机科学与博弈论 · 计算机科学 2024-08-07 Thomas Colcombet , Nathanaël Fijalkow , Florian Horn

We study quantum algorithms for spatial search on finite dimensional grids. Patel et al. and Falk have proposed algorithms based on a quantum walk without a coin, with different operators applied at even and odd steps. Until now, such…

量子物理 · 物理学 2015-10-14 Andris Ambainis , Renato Portugal , Nikolay Nahimov

Many popular reinforcement learning problems (e.g., navigation in a maze, some Atari games, mountain car) are instances of the episodic setting under its stochastic shortest path (SSP) formulation, where an agent has to achieve a goal state…

机器学习 · 统计学 2020-08-18 Jean Tarbouriech , Evrard Garcelon , Michal Valko , Matteo Pirotta , Alessandro Lazaric

We consider concurrent games played on graphs. At every round of the game, each player simultaneously and independently selects a move; the moves jointly determine the transition to a successor state. Two basic objectives are the safety…

计算机科学与博弈论 · 计算机科学 2008-12-18 Krishnendu Chatterjee , Luca de Alfaro , Thomas A. Henzinger

We explore the case of a group of random walkers looking for a target randomly located in space, such that the number of walkers is not constant but new ones can join the search, or those that are active can abandon it, with constant rates…

统计力学 · 物理学 2023-11-30 Daniel Campos , Vicenç Méndez

Coordinated teamwork is essential in fast-paced decision-making environments that require dynamic adaptation, often without an opportunity for explicit communication. Although implicit coordination has been extensively considered in the…

人工智能 · 计算机科学 2025-09-12 Thuy Ngoc Nguyen , Anita Williams Woolley , Cleotilde Gonzalez

Moving Target Defense (MTD) is commonly formulated as a repeated security game to mitigate persistent threats. Although the strong Stackelberg equilibrium (SSE) characterizes the defender's optimal strategy in the leader-follower framework,…

计算机科学与博弈论 · 计算机科学 2026-05-11 Zhaoyang Cheng , Guanpu Chen , Yiguang Hong , Ming Cao , Mikael Skoglund

The problem of opportunistic spectrum access in cognitive radio networks has been recently formulated as a non-Bayesian restless multi-armed bandit problem. In this problem, there are N arms (corresponding to channels) and one player…

机器学习 · 计算机科学 2011-11-10 Wenhan Dai , Yi Gai , Bhaskar Krishnamachari

We consider concurrent games played on graphs. At every round of a game, each player simultaneously and independently selects a move; the moves jointly determine the transition to a successor state. Two basic objectives are the safety…

计算机科学与博弈论 · 计算机科学 2008-09-25 Krishnendu Chatterjee , Luca de Alfaro , Thomas A. Henzinger

We study security games in which a defender commits to a mixed strategy for protecting a finite set of targets of different values. An attacker, knowing the defender's strategy, chooses which target to attack and for how long. If the…

计算机科学与博弈论 · 计算机科学 2018-04-24 David Kempe , Leonard J. Schulman , Omer Tamuz

The Competing Bandits framework is a recently emerging area that integrates multi-armed bandits in online learning with stable matching in game theory. While conventional models assume that all players and arms are constantly available, in…

机器学习 · 计算机科学 2026-03-23 Shinnosuke Uba , Yutaro Yamaguchi

We study concurrent stochastic reachability games played on finite graphs. Two players, Max and Min, seek respectively to maximize and minimize the probability of reaching a set of target states. We prove that Max has a memoryless strategy…

计算机科学中的逻辑 · 计算机科学 2024-01-25 Stefan Kiefer , Richard Mayr , Mahsa Shirmohammadi , Patrick Totzke

We study a variant of the searching problem where the environment consists of a known terrain and the goal is to obtain visibility of an unknown target point on the surface of the terrain. The searcher starts on the surface of the terrain…

计算几何 · 计算机科学 2024-01-03 Sarita de Berg , Nathan van Beusekom , Max van Mulken , Kevin Verbeek , Jules Wulms

We consider the problem of exploration of an anonymous, port-labeled, undirected graph with $n$ nodes and $m$ edges and diameter $D$, by a single mobile agent. Initially the agent does not know the graph topology nor any of the global…

数据结构与算法 · 计算机科学 2015-12-01 Artur Menc , Dominik Pająk , Przemysław Uznański

We study the problem of searching for a target at some unknown location in $\mathbb{R}^d$ when additional information regarding the position of the target is available in the form of predictions. In our setting, predictions come as…

计算几何 · 计算机科学 2025-04-08 Sergio Cabello , Panos Giannopoulos

Assume that a target is known to be present at an unknown point among a finite set of locations in the plane. We search for it using a mobile robot that has imperfect sensing capabilities. It takes time for the robot to move between…

The window mechanism, introduced by Chatterjee et al. for mean-payoff and total-payoff objectives in two-player turn-based games on graphs, refines long-term objectives with time bounds. This mechanism has proven useful in a variety of…

计算机科学与博弈论 · 计算机科学 2022-05-10 James C. A. Main , Mickael Randour , Jeremy Sproston

In collective tree exploration, a team of $k$ mobile agents is tasked to go through all edges of an unknown tree as fast as possible. An edge of the tree is revealed to the team when one agent becomes adjacent to that edge. The agents start…

数据结构与算法 · 计算机科学 2023-11-01 Romain Cosson

In this paper we investigate a differential game in which countably many dynamical objects pursue a single one. All the players perform simple motions. The duration of the game is fixed. The controls of a group of pursuers are subject to…

最优化与控制 · 数学 2014-10-10 Mehdi Salimi , Gafurjan Ibragimov , Stefan Siegmund , Somayeh Sharifi