中文
相关论文

相关论文: Stackelberg Punishment and Bully-Proofing Autonomo…

200 篇论文

We present a new concept called Game Mechanic Alignment theory as a way to organize game mechanics through the lens of systemic rewards and agential motivations. By disentangling player and systemic influences, mechanics may be better…

人工智能 · 计算机科学 2021-08-12 Michael Cerny Green , Ahmed Khalifa , Philip Bontrager , Rodrigo Canaan , Julian Togelius

A large body of empirical evidence suggests that humans are willing to engage in costly punishment of defectors in public goods games. Based on such pieces of evidence, it is suggested that punishment serves an important role in promoting…

物理与社会 · 物理学 2022-01-25 Mohammad Salahshour

Interactions between pedestrians, bikers, and human-driven vehicles have been a major concern in traffic safety over the years. The upcoming age of autonomous vehicles will further raise major problems on whether self-driving cars can…

计算机科学与博弈论 · 计算机科学 2018-06-26 Umberto Michieli , Leonardo Badia

With high scalability, high video streaming quality, and low bandwidth requirement, peer-to-peer (P2P) systems have become a popular way to exchange files and deliver multimedia content over the internet. However, current P2P systems are…

网络与互联网体系结构 · 计算机科学 2014-08-05 Xin Kang , Yongdong Wu

This paper presents an optimization framework to model Formula 1 racing dynamics, where multiple cars interact physically and strategically. Aerodynamic wake effects, trajectory optimization, and energy management are integrated by means of…

To improve efficiency and reduce failures in autonomous vehicles, research has focused on developing robust and safe learning methods that take into account disturbances in the environment. Existing literature in robust reinforcement…

机器学习 · 计算机科学 2019-03-12 Xiaobai Ma , Katherine Driggs-Campbell , Mykel J. Kochenderfer

We study multi-player general-sum Markov games with one of the players designated as the leader and the other players regarded as followers. In particular, we focus on the class of games where the followers are myopic, i.e., they aim to…

机器学习 · 计算机科学 2021-12-28 Han Zhong , Zhuoran Yang , Zhaoran Wang , Michael I. Jordan

Internet tracking technologies and wearable electronics provide a vast amount of data to machine learning algorithms. This stock of data stands to increase with the developments of the internet of things and cyber-physical systems. Clearly,…

密码学与安全 · 计算机科学 2016-08-11 Jeffrey Pawlick , Quanyan Zhu

We consider a repeated sequential game between a learner, who plays first, and an opponent who responds to the chosen action. We seek to design strategies for the learner to successfully interact with the opponent. While most previous…

机器学习 · 计算机科学 2020-07-13 Pier Giuseppe Sessa , Ilija Bogunovic , Maryam Kamgarpour , Andreas Krause

We study a two-player dynamic Stackelberg game where the follower's intention is unknown to the leader. Classical formulations of the Stackelberg equilibrium (SE) assume that the follower's best response (BR) function is known to the…

系统与控制 · 电气工程与系统科学 2026-04-09 Cayetana Salinas-Rodriguez , Jonathan Rogers , Sarah H. Q. Li

Automated decision-making tools increasingly assess individuals to determine if they qualify for high-stakes opportunities. A recent line of research investigates how strategic agents may respond to such scoring tools to receive favorable…

机器学习 · 计算机科学 2021-10-28 Keegan Harris , Hoda Heidari , Zhiwei Steven Wu

We initiate the study of structured Stackelberg games, a novel form of strategic interaction between a leader and a follower where contextual information can be predictive of the follower's (unknown) type. Motivated by applications such as…

计算机科学与博弈论 · 计算机科学 2026-05-18 Maria-Florina Balcan , Kiriaki Fragkia , Keegan Harris

We address two-player general-sum stochastic Stackelberg games (SSGs), where the leader's policy is optimized considering the best-response follower whose policy is optimal for its reward under the leader. Existing policy gradient and value…

计算机科学与博弈论 · 计算机科学 2026-03-17 Mikoto Kudo , Youhei Akimoto

The collective of autonomous cars is expected to generate almost optimal traffic. In this position paper we discuss the multi-agent models and the verification results of the collective behaviour of autonomous cars. We argue that…

多智能体系统 · 计算机科学 2017-09-11 László Z. Varga

We propose a single-level numerical approach to solve Stackelberg mean field game (MFG) problems. In Stackelberg MFG, an infinite population of agents play a non-cooperative game and choose their controls to optimize their individual…

最优化与控制 · 数学 2024-04-24 Gokce Dayanikli , Mathieu Lauriere

We study a repeated game between a supplier and a retailer who want to maximize their respective profits without full knowledge of the problem parameters. After characterizing the uniqueness of the Stackelberg equilibrium of the stage game…

计算机科学与博弈论 · 计算机科学 2022-07-12 Nicolò Cesa-Bianchi , Tommaso Cesari , Takayuki Osogami , Marco Scarsini , Segev Wasserkrug

Costly punishment has been suggested as a key mechanism for stabilizing cooperation in one-shot games. However, recent studies have revealed that the effectiveness of costly punishment can be diminished by second-order free riders (i.e.,…

物理与社会 · 物理学 2024-03-19 Chen Shen , Zhixue He , Lei Shi , Zhen Wang , Jun Tanimoto

We consider a two-player zero-sum network routing game in which a router wants to maximize the amount of legitimate traffic that flows from a given source node to a destination node and an attacker wants to block as much legitimate traffic…

计算机科学与博弈论 · 计算机科学 2020-03-13 David Grimsman , Joao P Hespanha , Jason R Marden

This paper considers the problem of how to allocate power among competing users sharing a frequency-selective interference channel. We model the interaction between selfish users as a non-cooperative game. As opposed to the existing…

计算机科学与博弈论 · 计算机科学 2008-12-16 Yi Su , Mihaela van der Schaar

When deployed in the world, a learning agent such as a recommender system or a chatbot often repeatedly interacts with another learning agent (such as a user) over time. In many such two-agent systems, each agent learns separately and the…

机器学习 · 计算机科学 2024-06-24 Kate Donahue , Nicole Immorlica , Meena Jagadeesan , Brendan Lucier , Aleksandrs Slivkins