中文
相关论文

相关论文: Parametric Bounded L\"ob's Theorem and Robust Coop…

200 篇论文

According to Strachey, a polymorphic program is parametric if it applies a uniform algorithm independently of the type instantiations at which it is applied. The notion of relational parametricity, introduced by Reynolds, is one possible…

编程语言 · 计算机科学 2019-03-14 Rasmus Ejlers Møgelberg , Alex Simpson

In experimental applications of bounded-reasoning models, behavior is often summarized by distributions of "levels". We argue that such summaries conflate two conceptually distinct dimensions: a player's type, capturing beliefs about what…

理论经济学 · 经济学 2026-04-15 Shuige Liu , Gabriel Ziegler

We present a framework that incorporates the idea of bounded rationality into dynamic stochastic pursuit-evasion games. The solution of a stochastic game is characterized, in general, by its (Nash) equilibria in feedback form. However,…

系统与控制 · 电气工程与系统科学 2020-03-17 Yue Guan , Dipankar Maity , Christopher M. Kroninger , Panagiotis Tsiotras

This work considers coordination and bargaining between two selfish users over a Gaussian interference channel. The usual information theoretic approach assumes full cooperation among users for codebook and rate selection. In the scenario…

信息论 · 计算机科学 2015-03-17 Xi Liu , Elza Erkip

Model-based algorithms -- algorithms that explore the environment through building and utilizing an estimated model -- are widely used in reinforcement learning practice and theoretically shown to achieve optimal sample efficiency for…

机器学习 · 计算机科学 2021-02-09 Qinghua Liu , Tiancheng Yu , Yu Bai , Chi Jin

We consider 2-player games played on a finite state space for infinite rounds. The games are concurrent: in each round, the two players choose their moves simultaneously; the current state and the moves determine the successor. We consider…

计算机科学与博弈论 · 计算机科学 2013-06-21 Krishnendu Chatterjee

As Large Language Models (LLMs) are integrated into critical real-world applications, their strategic and logical reasoning abilities are increasingly crucial. This paper evaluates LLMs' reasoning abilities in competitive environments…

Cooperation through repetition is an important theme in game theory. In this regard, various celebrated ``folk theorems'' have been proposed for repeated games in increasingly more complex environments. There has, however, been insufficient…

理论经济学 · 经济学 2024-02-16 Richard McLean , Ichiro Obara , Andrew Postlewaite

We consider a class of nonautonomous parabolic first-order coupled systems in the Lebesgue space $L^p({\mathbb R}^d;{\mathbb R}^m)$, $(d,m \ge 1)$ with $p\in [1,+\infty)$. Sufficient conditions for the associated evolution operator ${\bf…

偏微分方程分析 · 数学 2015-05-20 Luciana Angiuli , Luca Lorenzi , Diego Pallara

Multi-agent learning is a challenging problem in machine learning that has applications in different domains such as distributed control, robotics, and economics. We develop a prescriptive model of multi-agent behavior using Markov games.…

人工智能 · 计算机科学 2020-05-27 Jalal Etesami , Christoph-Nikolas Straehle

Exploiting others is beneficial individually but it could also be detrimental globally. The reverse is also true: a higher cooperation level may change the environment in a way that is beneficial for all competitors. To explore the possible…

物理与社会 · 物理学 2018-02-23 Attila Szolnoki , Xiaojie Chen

Agentic theorem provers combine a reasoning model, retrieval, search, and a proof assistant verifier, yet it remains unclear which components actually improve finite-budget proof success and why they help on real mathematical workloads. We…

机器学习 · 统计学 2026-05-26 Sho Sonoda , Shunta Akiyama , Yuya Uezato

The Prisoner's Dilemma Process on a graph $G$ is an iterative process where each vertex, with a fixed strategy (cooperate or defect), plays the game with each of its neighbours. At the end of a round each vertex may change its strategy to…

离散数学 · 计算机科学 2017-11-22 Christopher Duffy , Jeannette Janssen

We study multi-objective reinforcement learning (RL) where an agent's reward is represented as a vector. In settings where an agent competes against opponents, its performance is measured by the distance of its average return vector to a…

机器学习 · 计算机科学 2021-02-08 Tiancheng Yu , Yi Tian , Jingzhao Zhang , Suvrit Sra

The dominant theories of rational choice assume logical omniscience. That is, they assume that when facing a decision problem, an agent can perform all relevant computations and determine the truth value of all relevant logical/mathematical…

人工智能 · 计算机科学 2023-07-12 Caspar Oesterheld , Abram Demski , Vincent Conitzer

Partitioning a large group of employees into teams can prove difficult because unsatisfied employees may want to transfer to other teams. In this case, the team (coalition) formation is unstable and incentivizes deviation from the proposed…

计算机科学与博弈论 · 计算机科学 2024-06-04 Martin Bullinger , Sonja Kraiczy

We consider a group of agents on a graph who repeatedly play the prisoner's dilemma game against their neighbors. The players adapt their actions to the past behavior of their opponents by applying the win-stay lose-shift strategy. On a…

概率论 · 数学 2007-05-23 Elchanan Mossel , Sebastien Roch

Large scale systems are forecasted to greatly impact our future lives thanks to their wide ranging applications including cooperative robotics, mobility on demand, resource allocation, supply chain management. While technological…

最优化与控制 · 数学 2024-12-20 Dario Paccagnan

The overall aim of our research is to develop techniques to reason about the equilibrium properties of multi-agent systems. We model multi-agent systems as concurrent games, in which each player is a process that is assumed to act…

计算机科学中的逻辑 · 计算机科学 2020-08-14 Julian Gutierrez , Aniello Murano , Giuseppe Perelli , Sasha Rubin , Thomas Steeples , Michael Wooldridge

Coding theory plays a crucial role in ensuring data integrity and reliability across various domains, from communication to computation and storage systems. However, its reliance on trust assumptions for data recovery, which requires the…

信息论 · 计算机科学 2026-01-15 Hanzaleh Akbari Nodehi , Viveck R. Cadambe , Mohammad Ali Maddah-Ali