中文
相关论文

相关论文: Anchoring Theory in Sequential Stackelberg Games

200 篇论文

While Nash equilibria are guaranteed to exist, they may exhibit dense support, making them difficult to understand and execute in some applications. In this paper, we study $k$-sparse commitments in games where one player is restricted to…

计算机科学与博弈论 · 计算机科学 2025-04-22 Salam Afiouni , Jakub Černý , Chun Kai Ling , Christian Kroer

Mixed-integer linear programming (MILP) has been a fundamental problem in combinatorial optimization. Conventional MILP solving mainly relies on carefully designed heuristics embedded in the branch-and-bound framework. Driven by the strong…

人工智能 · 计算机科学 2026-01-13 Siyuan Li , Yifan Yu , Zhihao Zhang , Mengjing Chen , Fangzhou Zhu , Tao Zhong , Peng Liu , Jianye Hao

In general-sum games, the interaction of self-interested learning agents commonly leads to socially worse outcomes, such as defect-defect in the iterated stag hunt (ISH). Previous works address this challenge by sharing rewards or shaping…

多智能体系统 · 计算机科学 2023-03-15 Ziyi Liu , Yongchun Fang

Macroeconomic outcomes emerge from individuals' decisions, making it essential to model how agents interact with macro policy via consumption, investment, and labor choices. We formulate this as a dynamic Stackelberg game: the government…

理论经济学 · 经济学 2025-06-03 Qirui Mi , Zhiyu Zhao , Chengdong Ma , Siyu Xia , Yan Song , Mengyue Yang , Jun Wang , Haifeng Zhang

The role of uncertainty in data management has become more prominent than ever before, especially because of the growing importance of machine learning-driven applications that produce large uncertain databases. A well-known approach to…

数据库 · 计算机科学 2023-04-13 Efthymia Tsamoura , Jaehun Lee , Jacopo Urbani

This paper considers the discounted criterion of nonzero-sum decentralized stochastic games with prospect players. The state and action spaces are finite. The state transition probability is nonstationary. Each player independently controls…

最优化与控制 · 数学 2024-05-16 Yiting Wu , Junyu Zhang

This paper is devoted to a Stackelberg stochastic differential game for a linear mean-field type stochastic differential system with a mean-field type quadratic cost functional in finite horizon. The coefficients in the state equation and…

最优化与控制 · 数学 2023-08-22 Zixuan Li , Jingtao Shi

Bounded rationality refers to the non-optimal rationality of players in non-cooperative games. In a networked game, the bounded rationality of players may be heterogeneous and spatially distributed. It has been shown that the `system…

适应与自组织系统 · 物理学 2022-03-25 Prasan Ratnayake , Dharshana Kasthurirathna , Mahendra Piraveenan

We introduce a novel framework for computing optimal randomized security policies in networked domains which extends previous approaches in several ways. First, we extend previous linear programming techniques for Stackelberg security games…

计算机科学与博弈论 · 计算机科学 2012-10-19 Joshua Letchford , Yevgeniy Vorobeychik

Algorithms for playing in Stackelberg games have been deployed in real-world domains including airport security, anti-poaching efforts, and cyber-crime prevention. However, these algorithms often fail to take into consideration the…

计算机科学与博弈论 · 计算机科学 2024-10-15 Keegan Harris , Zhiwei Steven Wu , Maria-Florina Balcan

Several blockchain consensus protocols proposed to use of Directed Acyclic Graphs (DAGs) to solve the limited processing throughput of traditional single-chain Proof-of-Work (PoW) blockchains. Many such protocols utilize a random…

密码学与安全 · 计算机科学 2023-05-29 Martin Perešíni , Ivan Homoliak , Federico Matteo Benčić , Martin Hrubý , Kamil Malinka

In this paper, we consider a discrete-time stochastic Stackelberg game with a single leader and multiple followers. Both the followers and the leader together have conditionally independent private types, conditioned on action and previous…

最优化与控制 · 数学 2022-09-21 Deepanshu Vasal

Machine learning relies on the assumption that unseen test instances of a classification problem follow the same distribution as observed training data. However, this principle can break down when machine learning is used to make important…

机器学习 · 计算机科学 2015-11-24 Moritz Hardt , Nimrod Megiddo , Christos Papadimitriou , Mary Wootters

Hierarchical reasoning model (HRM) achieves extraordinary performance on various reasoning tasks, significantly outperforming large language model-based reasoners. To understand the strengths and potential failure modes of HRM, we conduct a…

人工智能 · 计算机科学 2026-03-24 Zirui Ren , Ziming Liu

Dealing with context dependent knowledge has led to different formalizations of the notion of context. Among them is the Contextualized Knowledge Repository (CKR) framework, which is rooted in description logics but links on the reasoning…

人工智能 · 计算机科学 2021-12-23 Loris Bozzato , Thomas Eiter , Rafael Kiesel

This thesis develops theoretical frameworks and algorithms that advance constrained reinforcement learning (RL) across control, preference learning, and alignment of large language models. The first contribution addresses constrained Markov…

机器学习 · 计算机科学 2025-12-12 Akhil Agnihotri

Answer set programming (ASP) is a paradigm for declarative problem solving where problems are first formalized as rule sets, i.e., answer-set programs, in a uniform way and then solved by computing answer sets for programs. The…

人工智能 · 计算机科学 2011-08-31 Mai Nguyen , Tomi Janhunen , Ilkka Niemelä

Because an agents resources dictate what actions it can possibly take, it should plan which resources it holds over time carefully, considering its inherent limitations (such as power or payload restrictions), the competing needs of other…

多智能体系统 · 计算机科学 2014-01-17 Jianhui Wu , Edmund H. Durfee

There is a growing interest in studying sequential neural posterior estimation (SNPE) techniques due to their advantages for simulation-based models with intractable likelihoods. The methods aim to learn the posterior from adaptively…

统计计算 · 统计学 2025-10-16 Xiliang Yang , Yifei Xiong , Zhijian He

In this work, we provide a structural characterization of the possible Nash equilibria in the well-studied class of security games with additive utility. Our analysis yields a classification of possible equilibria into seven types and we…

计算机科学与博弈论 · 计算机科学 2022-08-05 Joe Clanin , Sourabh Bhattacharya