中文
相关论文

相关论文: Integrating Sequential Hypothesis Testing into Adv…

200 篇论文

We propose an extension of Strategy Logic (SL), in which one can both reason about strategizing under imperfect information and about players' knowledge. One original aspect of our approach is that we do not force strategies to be uniform,…

计算机科学中的逻辑 · 计算机科学 2019-08-08 Sophia Knight , Bastien Maubert

The emergence of cooperation in the thermodynamic limit of social dilemmas is an emerging field of research. While numerical approaches (using replicator dynamics) are dime a dozen, analytical approaches are rare. A particularly useful…

物理与社会 · 物理学 2020-09-10 Colin Benjamin , Aditya Dash

We investigate the linear quadratic Gaussian Stackelberg game under a class of nested observation information pattern. Two decision makers implement control strategies relying on different information sets: The follower uses its observation…

最优化与控制 · 数学 2022-06-07 Zhipeng Li , Damian Marelli , Minyue Fu , Huanshui Zhang

Selection of input features such as relevant pieces of text has become a common technique of highlighting how complex neural predictors operate. The selection can be optimized post-hoc for trained models or incorporated directly into the…

机器学习 · 计算机科学 2019-10-29 Shiyu Chang , Yang Zhang , Mo Yu , Tommi S. Jaakkola

A fundamental challenge in formal theorem proving by LLMs is the lack of high-quality training data. Although reinforcement learning or expert iteration partially mitigates this issue by alternating between LLM generating proofs and…

机器学习 · 计算机科学 2025-03-24 Kefan Dong , Tengyu Ma

Game theory provides the gold standard for analyzing adversarial engagements, offering strong optimality guarantees. However, these guarantees often become brittle when assumptions such as perfect information are violated. Reinforcement…

机器学习 · 计算机科学 2026-03-18 Goutam Das , Michael Dorothy , Kyle Volle , Daigo Shishika

Continuous post-deployment compliance audits, mandated by emerging regulations such as the EU AI Act and Digital Services Act, create a class of strategic gaming distinct from the one-shot input/output gaming studied in prior work.…

计算机与社会 · 计算机科学 2026-05-08 Florian A. D. Burnat , Brittany I. Davidson

Many situations require people to act quickly and are characterized by asymmetric information. Since asymmetric information makes people tempted to misreport their private information for their own benefit, it is of primary importance to…

物理与社会 · 物理学 2017-05-05 Valerio Capraro

Betting games provide a natural setting to capture how information yields strategic advantage. The Kelly criterion for betting, long a cornerstone of portfolio theory and information theory, admits an interpretation in the limit of…

量子物理 · 物理学 2026-01-15 Maite Arcos , Renato Renner , Jonathan Oppenheim

In this paper, we investigate a new model of a linear-quadratic mean-field stochastic Stackelberg differential game with one leader and two followers, in which the leader is allowed to stop her strategy at a random time. Our overarching…

最优化与控制 · 数学 2021-06-08 Zhun Gou , Nan-jing Huang , Ming-hui Wang

Federated learning is a distributed learning paradigm where multiple agents, each only with access to local data, jointly learn a global model. There has recently been an explosion of research aiming not only to improve the accuracy rates…

计算机科学与博弈论 · 计算机科学 2021-06-18 Kate Donahue , Jon Kleinberg

To evaluate the safety and usefulness of deployment protocols for untrusted AIs, AI Control uses a red-teaming exercise played between a protocol designer and an adversary. This paper introduces AI-Control Games, a formal decision-making…

人工智能 · 计算机科学 2026-05-08 Charlie Griffin , Louis Thomson , Buck Shlegeris , Alessandro Abate

In many multi-agent systems, agents interact repeatedly and are expected to settle into stable, rational behavior over time. Yet in practice, behavior often drifts, and detecting such deviations in real time remains an open challenge. We…

计算机科学与博弈论 · 计算机科学 2026-05-25 Etienne Gauthier , Francis Bach , Michael I. Jordan

The exponential growth of data volumes has led to escalating computational costs in machine learning model training. However, many features fail to contribute positively to model performance while consuming substantial computational…

机器学习 · 计算机科学 2025-12-01 Chi Zhao , Jing Liu , Elena Parilina

We study a two-player dynamic Stackelberg game where the follower's intention is unknown to the leader. Classical formulations of the Stackelberg equilibrium (SE) assume that the follower's best response (BR) function is known to the…

系统与控制 · 电气工程与系统科学 2026-04-09 Cayetana Salinas-Rodriguez , Jonathan Rogers , Sarah H. Q. Li

As assembly tasks grow in complexity, collaboration among multiple robots becomes essential for task completion. However, centralized task planning has become inadequate for adapting to the increasing intelligence and versatility of robots,…

机器人学 · 计算机科学 2024-04-22 Yuhan Zhao , Lan Shi , Quanyan Zhu

The Iterated Prisoner's Dilemma has guided research on social dilemmas for decades. However, it distinguishes between only two atomic actions: cooperate and defect. In real-world prisoner's dilemmas, these choices are temporally extended…

人工智能 · 计算机科学 2018-03-02 Weixun Wang , Jianye Hao , Yixi Wang , Matthew Taylor

We present an approach for systematically anticipating the actions and policies employed by \emph{oblivious} environments in concurrent stochastic games, while maximizing a reward function. Our main contribution lies in the synthesis of a…

人工智能 · 计算机科学 2024-09-19 Shadi Tasdighi Kalat , Sriram Sankaranarayanan , Ashutosh Trivedi

In social dilemma situations, individual rationality leads to sub-optimal group outcomes. Several human engagements can be modeled as a sequential (multi-step) social dilemmas. However, in contrast to humans, Deep Reinforcement Learning…

It is commonly assumed that trust increases cooperation. However, game-theoretic models often fail to distinguish between cooperative actions and trust, making it difficult to independently measure trust and determine how its effects vary…

计算机科学与博弈论 · 计算机科学 2026-03-18 Cedric Perret , The Anh Han , Elias Fernández Domingos , Theodor Cimpeanu , Simon T. Powers
‹ 上一页 1 8 9 10 下一页 ›