中文
相关论文

相关论文: Stochastic Perron's method and elementary strategi…

200 篇论文

We consider zero-sum stochastic games with perfect information and finitely many states and actions. The payoff is computed by a function which associates to each infinite sequence of states and actions a real number. We prove that if the…

计算机科学与博弈论 · 计算机科学 2022-03-29 Hugo Gimbert , Edon Kelmendi

We consider discrete time partially observable zero-sum stochastic game with average payoff criterion. We study the game using an equivalent completely observable game. We show that the game has a value and also we come up with a pair of…

最优化与控制 · 数学 2014-09-16 Subhamay Saha

We consider a stochastic differential equation that is controlled by means of an additive finite-variation process. A singular stochastic controller, who is a minimizer, determines this finite-variation process, while a discretionary…

概率论 · 数学 2015-01-20 Daniel Hernandez-Hernandez , Robert S. Simon , Mihail Zervos

We consider a zero-sum stochastic game for continuous-time Markov chain with countable state space and unbounded transition and pay-off rates. The additional feature of the game is that the controllers together with taking actions are also…

最优化与控制 · 数学 2020-09-01 Chandan Pal , Subhamay Saha

Decoding strategies play a pivotal role in text generation for modern language models, yet a puzzling gap divides theory and practice. Surprisingly, strategies that should intuitively be optimal, such as Maximum a Posteriori (MAP), often…

机器学习 · 计算机科学 2025-05-20 Sijin Chen , Omar Hagrass , Jason M. Klusowski

We study zero-sum differential games with state constraints and one-sided information, where the informed player (Player 1) has a categorical payoff type unknown to the uninformed player (Player 2). The goal of Player 1 is to minimize his…

计算机科学与博弈论 · 计算机科学 2024-06-05 Mukesh Ghimire , Lei Zhang , Zhe Xu , Yi Ren

The purpose of this paper is to study 2-person zero-sum stochastic differential games, in which one player is a major one and the other player is a group of $N$ minor agents which are collectively playing, statistically identical and have…

概率论 · 数学 2013-08-26 Rainer Buckdahn , Juan Li , Shige Peng

We present a novel framework for {\epsilon}-optimally solving two-player zero-sum partially observable stochastic games (zs-POSGs). These games pose a major challenge due to the absence of a principled connection with dynamic programming…

计算机科学与博弈论 · 计算机科学 2025-11-17 Erwan Christian Escudie , Matthia Sabatelli , Olivier Buffet , Jilles Steeve Dibangoye

This paper is concerned with a new type of differential game problems of forwardbackward stochastic systems. There are three distinguishing features: Firstly, our game systems are forward-backward doubly stochastic differential equations,…

最优化与控制 · 数学 2015-10-09 Eddie C. M. Hui , Hua Xiao

We formulate a new class of two-person zero-sum differential games, in a stochastic setting, where a specification on a target terminal state distribution is imposed on the players. We address such added specification by introducing…

系统与控制 · 电气工程与系统科学 2019-09-13 Yongxin Chen , Tryphon T. Georgiou , Michele Pavon

We consider a game, in which the dynamics is described by a non-linear Volterra integral equation of Hammerstein type with a weakly-singular kernel and the goals of the first and second players are, respectively, to minimize and maximize a…

最优化与控制 · 数学 2024-04-17 Mikhail I. Gomoyunov

Stochastic games are an important class of problems that generalize Markov decision processes to game theoretic scenarios. We consider finite state two-player zero-sum stochastic games over an infinite time horizon with discounted rewards.…

最优化与控制 · 数学 2008-06-17 Parikshit Shah , Pablo A. Parrilo

We present a novel variant of fictitious play dynamics combining classical fictitious play with Q-learning for stochastic games and analyze its convergence properties in two-player zero-sum stochastic games. Our dynamics involves players…

计算机科学与博弈论 · 计算机科学 2022-06-03 Muhammed O. Sayin , Francesca Parise , Asuman Ozdaglar

Through a stochastic control theoretic approach, we analyze reputation games where a strategic long-lived player acts in a sequential repeated game against a collection of short-lived players. The key assumption in our model is that the…

最优化与控制 · 数学 2020-01-22 Nuh Aygün Dalkıran , Serdar Yüksel

We consider a sub-class of bi-matrix games which we refer to as two-person (hereafter referred to as two-player) additively-separable sum (TPASS) games, where the sum of the pay-offs of the two players is additively separable. The row…

最优化与控制 · 数学 2025-11-03 Somdeb Lahiri

Recent results of Ye and Hansen, Miltersen and Zwick show that policy iteration for one or two player (perfect information) zero-sum stochastic games, restricted to instances with a fixed discount rate, is strongly polynomial. We show that…

最优化与控制 · 数学 2013-10-21 Marianne Akian , Stéphane Gaubert

Motivated by a vaccination coverage problem, we consider here a zero-sum differential game governed by a differential system consisting of a hyperbolic partial differential equation (PDE) and an ordinary differential equation (ODE). Two…

偏微分方程分析 · 数学 2024-12-18 Mauro Garavello , Elena Rossi , Abraham Sylla

Driven by recent successes in two-player, zero-sum game solving and playing, artificial intelligence work on games has increasingly focused on algorithms that produce equilibrium-based strategies. However, this approach has been less…

计算机科学与博弈论 · 计算机科学 2022-06-24 Dustin Morrill , Ryan D'Orazio , Reca Sarfati , Marc Lanctot , James R. Wright , Amy Greenwald , Michael Bowling

We consider an N-player hierarchical game in which the i-th player's objective comprises of an expectation-valued term, parametrized by rival decisions, and a hierarchical term. Such a framework allows for capturing a broad range of…

最优化与控制 · 数学 2024-01-26 Shisheng Cui , Uday V. Shanbhag , Mathias Staudigl

We consider the general model of zero-sum repeated games (or stochastic games with signals), and assume that one of the players is fully informed and controls the transitions of the state variable. We prove the existence of the uniform…

最优化与控制 · 数学 2009-04-20 Jérôme Renault