中文
相关论文

相关论文: Solving Structured Hierarchical Games Using Differ…

200 篇论文

This article extends the idea of solving parity games by strategy iteration to non-deterministic strategies: In a non-deterministic strategy a player restricts himself to some non-empty subset of possible actions at a given node, instead of…

计算机科学与博弈论 · 计算机科学 2012-03-20 Michael Luttenberger

We study the performance of the gradient play algorithm for stochastic games (SGs), where each agent tries to maximize its own total discounted reward by making decisions independently based on current state information which is shared…

机器学习 · 计算机科学 2023-12-08 Runyu Zhang , Zhaolin Ren , Na Li

Stochastic differential games have been used extensively to model agents' competitions in Finance, for instance, in P2P lending platforms from the Fintech industry, the banking system for systemic risk, and insurance markets. The recently…

最优化与控制 · 数学 2021-03-23 Jiequn Han , Ruimeng Hu , Jihao Long

Dynamic games are powerful tools to model multi-agent decision-making, yet computing Nash (generalized Nash) equilibria remains a central challenge in such settings. Complexity arises from tightly coupled optimality conditions, nested…

计算机科学与博弈论 · 计算机科学 2026-02-06 Mahdis Rabbani , Navid Mojahed , Shima Nazari

Tree ensemble algorithms as RandomForest and GradientBoosting are currently the dominant methods for modeling discrete or tabular data, however, they are unable to perform a hierarchical representation learning from raw data as…

A general theory of stochastic decision forests is developed to bridge two concepts of information flow: decision trees and refined partitions on the one side, filtrations from probability theory on the other. Instead of the traditional…

理论经济学 · 经济学 2024-11-12 E. Emanuel Rapsch

Cooperation in heterogeneous groups, where individuals differ in resources, productivity, and behavioural responsiveness, underpins collective action across many social and biological systems. Introspection dynamics, in which each player…

计算机科学与博弈论 · 计算机科学 2026-05-25 Harry Foster , Vincent A. Knight , Sebastian Krapohl

In many settings of interest, a policy is set by one party, the leader, in order to influence the action of another party, the follower, where the follower's response is determined by some private information. A natural question to ask is,…

计算机科学与博弈论 · 计算机科学 2025-04-23 Michael Albert , Quinlan Dawkins , Minbiao Han , Haifeng Xu

In this paper, we establish a dynamic game to allocate CSR (Corporate Social Responsibility) to the members of a supply chain. We propose a model of three-tier supply chain in decentralized state that is including supplier, manufacturer and…

最优化与控制 · 数学 2015-03-17 Mehrnoosh Khademi , Massimiliano Ferrara , Bruno Pansera , Mehdi Salimi

We propose projection-free sequential algorithms for linear-quadratic dynamics games. These policy gradient based algorithms are akin to Stackelberg leadership model and can be extended to model-free settings. We show that if the leader…

系统与控制 · 电气工程与系统科学 2019-11-13 Jingjing Bu , Lillian J. Ratliff , Mehran Mesbahi

Game-theoretic resource allocation on graphs (GRAG) involves two players competing over multiple steps to control nodes of interest on a graph, a problem modeled as a multi-step Colonel Blotto Game (MCBG). Finding optimal strategies is…

机器学习 · 计算机科学 2025-05-13 Zijian An , Lifeng Zhou

We consider finite-horizon and infinite-horizon versions of a dynamic game with $N$ selfish players who observe their types privately and take actions that are publicly observed. Players' types evolve as conditionally independent Markov…

最优化与控制 · 数学 2018-03-20 Deepanshu Vasal , Abhinav Sinha , Achilleas Anastasopoulos

In a sequential decision-making problem, the information structure is the description of how events in the system occurring at different points in time affect each other. Classical models of reinforcement learning (e.g., MDPs, POMDPs)…

机器学习 · 计算机科学 2024-05-29 Awni Altabaa , Zhuoran Yang

In this paper, we establish a dynamic game to allocate CSR (Corporate Social Responsibility) to the members of a supply chain. We propose a model of a three-tier supply chain in a decentralized state which includes a supplier, a…

最优化与控制 · 数学 2015-06-23 Mehrnoosh Khademi , Massimiliano Ferrara , Mehdi Salimi , Somayeh Sharifi

Large language model (LLM) agents have shown remarkable progress in social deduction games (SDGs). However, existing approaches primarily focus on information processing and strategy selection, overlooking the significance of persuasive…

人工智能 · 计算机科学 2026-04-15 Zhang Zheng , Deheng Ye , Peilin Zhao , Hao Wang

Humans can leverage hierarchical structures to split a task into sub-tasks and solve problems efficiently. Both imitation and reinforcement learning or a combination of them with hierarchical structures have been proven to be an efficient…

机器人学 · 计算机科学 2020-12-15 Yaru Niu , Yijun Gu

In many settings where multiple agents interact, the optimal choices for each agent depend heavily on the choices of the others. These coupled interactions are well-described by a general-sum differential game, in which players have…

机器人学 · 计算机科学 2020-05-07 Lasse Peters , David Fridovich-Keil , Claire J. Tomlin , Zachary N. Sunberg

If the influence diagram (ID) depicting a Bayesian game is common knowledge to its players then additional assumptions may allow the players to make use of its embodied irrelevance statements. They can then use these to discover a simpler…

计算机科学与博弈论 · 计算机科学 2017-04-10 Peter A. Thwaites , Jim Q. Smith

Deep reinforcement learning for high dimensional, hierarchical control tasks usually requires the use of complex neural networks as functional approximators, which can lead to inefficiency, instability and even divergence in the training…

机器学习 · 计算机科学 2019-11-26 Yuguang Yang

In this study, we explore the application of game theory, in particular Stackelberg games, to address the issue of effective coordination strategy generation for heterogeneous robots with one-way communication. To that end, focusing on the…

机器人学 · 计算机科学 2023-08-01 Yuhan Zhao , Baichuan Huang , Jingjin Yu , Quanyan Zhu