English
Related papers

Related papers: Online Competitive Information Gathering for Parti…

200 papers

Partially observable Markov decision processes (POMDPs) are a general framework for sequential decision-making under latent state uncertainty, yet learning in POMDPs is intractable in the worst case. Motivated by sensing and probing…

Machine Learning · Computer Science 2026-01-27 Ming Shi , Yingbin Liang , Ness B. Shroff

Zero-sum stochastic games provide a rich model for competitive decision making. However, under general forms of state uncertainty as considered in the Partially Observable Stochastic Game (POSG), such decision making problems are still not…

Artificial Intelligence · Computer Science 2016-06-23 Auke J. Wiggers , Frans A. Oliehoek , Diederik M. Roijers

We consider multi-player graph games with partial-observation and parity objective. While the decision problem for three-player games with a coalition of the first and second players against the third player is undecidable, we present a…

Logic in Computer Science · Computer Science 2014-04-23 Krishnendu Chatterjee , Laurent Doyen

Formulating cyber-security problems with attackers and defenders as a partially observable stochastic game has become a trend recently. Among them, the one-sided two-player zero-sum partially observable stochastic game (OTZ-POSG) has…

Systems and Control · Electrical Eng. & Systems 2021-09-20 Wei Zheng , Taeho Jung , Hai Lin

We propose a stochastic first-order algorithm to learn the rationality parameters of simultaneous and non-cooperative potential games, i.e., the parameters of the agents' optimization problems. Our technique combines (i.) an active-set step…

Optimization and Control · Mathematics 2023-07-31 Stefan Clarke , Gabriele Dragotto , Jaime Fernández Fisac , Bartolomeo Stellato

This paper considers the distributed strategy design for Nash equilibrium (NE) seeking in multi-cluster games under a partial-decision information scenario. In the considered game, there are multiple clusters and each cluster consists of a…

Optimization and Control · Mathematics 2022-06-08 Min Meng , Xiuxian Li

Modeling the interaction between traffic agents is a key issue in designing safe and non-conservative maneuvers in autonomous driving. This problem can be challenging when multi-modality and behavioral uncertainties are engaged. Existing…

Robotics · Computer Science 2024-09-24 Zhenmin Huang , Tong Li , Shaojie Shen , Jun Ma

Estimating discrete games of complete information is often computationally difficult due to partial identification and the absence of closed-form moment characterizations. This paper proposes computationally tractable approaches to…

Econometrics · Economics 2025-10-02 Paul S. Koh

Distributed online optimization and game have been increasingly researched in the last decade, mostly motivated by its wide applications in sensor networks, robotics (e.g., distributed target tracking and formation control), smart grids,…

Machine Learning · Computer Science 2023-01-24 Xiuxian Li , Lihua Xie , Na Li

Search in test time is often used to improve the performance of reinforcement learning algorithms. Performing theoretically sound search in fully adversarial two-player games with imperfect information is notoriously difficult and requires…

Computer Science and Game Theory · Computer Science 2025-01-30 Ondrej Kubicek , Neil Burch , Viliam Lisy

We study a pursuit-evasion game between two players with car-like dynamics and sensing limitations by formalizing it as a partially observable stochastic zero-sum game. The partial observability caused by the sensing constraints is…

Robotics · Computer Science 2025-06-17 Burak M. Gonultas , Volkan Isler

Asymmetric information stochastic games (AISGs) arise in many complex socio-technical systems, such as cyber-physical systems and IT infrastructures. Existing computational methods for AISGs are primarily offline and can not adapt to…

Computer Science and Game Theory · Computer Science 2024-08-20 Tao Li , Kim Hammar , Rolf Stadler , Quanyan Zhu

We study synthesis problems with constraints in partially observable Markov decision processes (POMDPs), where the objective is to compute a strategy for an agent that is guaranteed to satisfy certain safety and performance specifications.…

This paper introduces an information-theoretic method for selecting a subset of problems which gives the most information about a group of problem-solving algorithms. This method was tested on the games in the General Video Game AI (GVGAI)…

Artificial Intelligence · Computer Science 2020-05-19 Matthew Stephenson , Damien Anderson , Ahmed Khalifa , John Levine , Jochen Renz , Julian Togelius , Christoph Salge

Cooperative games are those in which both agents share the same payoff structure. Value-based reinforcement-learning algorithms, such as variants of Q-learning, have been applied to learning cooperative games, but they only apply when the…

Machine Learning · Computer Science 2017-05-25 Leonid Peshkin , Kee-Eung Kim , Nicolas Meuleau , Leslie Pack Kaelbling

Cooperative games are those in which both agents share the same payoff structure. Value-based reinforcement-learning algorithms, such as variants of Q-learning, have been applied to learning cooperative games, but they only apply when the…

Artificial Intelligence · Computer Science 2014-08-08 Leonid Peshkin , Kee-Eung Kim , Nicolas Meuleau , Leslie Pack Kaelbling

Stochastic patrol routing is known to be advantageous in adversarial settings; however, the optimal choice of stochastic routing strategy is dependent on a model of the adversary. We adopt a worst-case omniscient adversary model from the…

Systems and Control · Electrical Eng. & Systems 2025-04-10 Yohan John , Gilberto Diaz-Garcia , Xiaoming Duan , Jason R. Marden , Francesco Bullo

We study \emph{partial-information} two-player turn-based games on graphs with omega-regular objectives, when the partial-information player has \emph{limited memory}. Such games are a natural formalization for reactive synthesis when the…

Formal Languages and Automata Theory · Computer Science 2020-02-19 Dhananjay Raju , Rüdiger Ehlers , Ufuk Topcu

Gaussian Process (GP) models are widely used for Robotic Information Gathering (RIG) in exploring unknown environments due to their ability to model complex phenomena with non-parametric flexibility and accurately quantify prediction…

Robotics · Computer Science 2024-06-07 Weizhe Chen , Lantao Liu , Roni Khardon

We study coalition formation in the framework of fractional hedonic games (FHGs). The objective is to maximize social welfare in an online model where agents arrive one by one and must be assigned to coalitions immediately and irrevocably.…

Computer Science and Game Theory · Computer Science 2025-10-22 Martin Bullinger , René Romen , Alexander Schlenga