中文
相关论文

相关论文: Fictitious Play with Time-Invariant Frequency Upda…

200 篇论文

Fictitious play (FP) is a natural learning dynamic in two-player zero-sum games. Samuel Karlin conjectured in 1959 that FP converges at a rate of $O(t^{-1/2})$ to Nash equilibrium, where $t$ is the number of steps played. However,…

计算机科学与博弈论 · 计算机科学 2025-07-15 Yuanhao Wang

While fictitious play is guaranteed to converge to Nash equilibrium in certain game classes, such as two-player zero-sum games, it is not guaranteed to converge in non-zero-sum and multiplayer games. We show that fictitious play in fact…

计算机科学与博弈论 · 计算机科学 2024-07-30 Sam Ganzfried

Fighting Fantasy is a popular recreational fantasy gaming system worldwide. Combat in this system progresses through a stochastic game involving a series of rounds, each of which may be won or lost. Each round, a limited resource (`luck')…

人工智能 · 计算机科学 2020-02-25 Iain G. Johnston

Continuous-time stochastic control with time-inhomogeneous jump-diffusion dynamics is central in finance and economics, but computing optimal policies is difficult under explicit time dependence, discontinuous shocks, and high…

最优化与控制 · 数学 2026-04-08 Liya Guo , Ruimeng Hu , Xu Yang , Yi Zhu

We consider the problem of distributed channel allocation in large networks under the frequency-selective interference channel. Performance is measured by the weighted sum of achievable rates. Our proposed algorithm is a modified Fictitious…

信息论 · 计算机科学 2018-11-13 Ilai Bistritz , Amir Leshem

Advanced persistent threat (APT) is a kind of stealthy, sophisticated, and long-term cyberattack that has brought severe financial losses and critical infrastructure damages. Existing works mainly focus on APT defense under stable network…

计算机科学与博弈论 · 计算机科学 2023-09-04 Zixuan Wang , Jiliang Li , Yuntao Wang , Zhou Su , Shui Yu , Weizhi Meng

We investigate a time-inconsistent, non-Markovian finite-player game in continuous time, where each player's objective functional depends non-linearly on the expected value of the state process. As a result, the classical Bellman optimality…

概率论 · 数学 2025-12-10 Dylan Possamaï , Chiara Rossato

We motivate and propose a new model for non-cooperative Markov game which considers the interactions of risk-aware players. This model characterizes the time-consistent dynamic "risk" from both stochastic state transitions (inherent to the…

计算机科学与博弈论 · 计算机科学 2019-11-22 Wenjie Huang , Pham Viet Hai , William B. Haskell

Experiments on the ultimatum game have revealed that humans are remarkably fond of fair play. When asked to share an amount of money, unfair offers are rare and their acceptance rate small. While empathy and spatiality may lead to the…

物理与社会 · 物理学 2012-08-20 Attila Szolnoki , Matjaz Perc , Gyorgy Szabo

This paper studies the problem of multi-step manipulative attacks in Stackelberg security games, in which a clever attacker attempts to orchestrate its attacks over multiple time steps to mislead the defender's learning of the attacker's…

人工智能 · 计算机科学 2022-03-02 Thanh H. Nguyen , Arunesh Sinha

Mean Field Game systems describe equilibrium configurations in differential games with infinitely many infinitesimal interacting agents. We introduce a learning procedure (similar to the Fictitious Play) for these games and show its…

最优化与控制 · 数学 2015-08-03 Pierre Cardaliaguet , Saeed Hadikhanloo

Considering the interaction through mutual interference of the different radio devices, the channel selection (CS) problem in decentralized parallel multiple access channels can be modeled by strategic-form games. Here, we show that the CS…

计算机科学与博弈论 · 计算机科学 2010-09-28 S. M. Perlaza , H. Tembine , S. Lasaulce , V. Quintero-Florez

A cyber security problem in a networked system formulated as a resilient graph problem based on a game-theoretic approach is considered. The connectivity of the underlying graph of the network system is reduced by an attacker who removes…

系统与控制 · 电气工程与系统科学 2023-03-14 Yurid Nugraha , Ahmet Cetinkaya , Tomohisa Hayakawa , Hideaki Ishii , Quanyan Zhu

We consider in discrete time, a general class of sequential stochastic dynamic games with asymmetric information with the following features. The underlying system has Markovian dynamics controlled by the agents' joint actions. Each agent's…

多智能体系统 · 计算机科学 2023-01-16 Yi Ouyang , Hamidreza Tavafoghi , Demosthenis Teneketzis

Sequential equilibrium is the conventional approach for analyzing multi-stage games of incomplete information. It relies on mutual consistency of beliefs. To relax mutual consistency, I theoretically and experimentally explore the dynamic…

理论经济学 · 经济学 2023-11-06 Po-Hsuan Lin

In the inference attacks studied in Quantitative Information Flow (QIF), the adversary typically tries to interfere with the system in the attempt to increase its leakage of secret information. The defender, on the other hand, typically…

密码学与安全 · 计算机科学 2023-07-19 Mário S. Alvim , Konstantinos Chatzikokolakis , Yusuke Kawamoto , Catuscia Palamidessi

In this work we consider a stochastic linear quadratic two-player game. The state measurements are observed through a switched noiseless communication link. Each player incurs a finite cost every time the link is established to get…

计算机科学与博弈论 · 计算机科学 2017-09-21 Dipankar Maity , Achilleas Anastasopoulos , John S. Baras

This paper addresses a mathematically tractable model of the Prisoner's Dilemma using the framework of active inference. In this work, we design pairs of Bayesian agents that are tracking the joint game state of their and their opponent's…

物理与社会 · 物理学 2023-08-31 Daphne Demekas , Conor Heins , Brennan Klein

Microscopic strategy update rules play an important role in the evolutionary dynamics of cooperation among interacting agents on complex networks. Many previous related works only consider one \emph{fixed} rule, while in the real world,…

最优化与控制 · 数学 2023-11-27 Shengxian Wang , Weijia Yao , Ming Cao , Xiaojie Chen

In stochastic games with incomplete information, the uncertainty is evoked by the lack of knowledge about a player's own and the other players' types, i.e. the utility function and the policy space, and also the inherent stochasticity of…

机器学习 · 计算机科学 2022-03-21 Hannes Eriksson , Debabrota Basu , Mina Alibeigi , Christos Dimitrakakis