中文
相关论文

相关论文: Application of the EWL protocol to decision proble…

200 篇论文

Learning in strategy games (e.g. StarCraft, poker) requires the discovery of diverse policies. This is often achieved by iteratively training new policies against existing ones, growing a policy population that is robust to exploit. This…

人工智能 · 计算机科学 2022-02-16 Siqi Liu , Luke Marris , Daniel Hennes , Josh Merel , Nicolas Heess , Thore Graepel

We consider the probabilistic planning problem where the agent (called Player 1, or P1) can jointly plan the control actions and sensor queries in a sensor network and an attacker (called player 2, or P2) can carry out attacks on the…

密码学与安全 · 计算机科学 2021-05-04 Abhishek N. Kulkarni , Shuo Han , Nandi O. Leslie , Charles A. Kamhoua , Jie Fu

Large Language Models (LLMs) reasoning abilities are increasingly being applied to classical board and card games, but the dominant approach -- involving prompting for direct move generation -- has significant drawbacks. It relies on the…

We study equivalence determination of unitary operations, a task analogous to quantum state discrimination. The candidate states are replaced by unitary operations given as a quantum sample, i.e., a black-box device implementing a candidate…

量子物理 · 物理学 2018-04-02 Atsushi Shimbo , Akihito Soeda , Mio Murao

In adversarial settings, a mobile agent may strategically plan its motion to influence an opponent's inference about its intended goal. We study deceptive path planning in a scenario where a mobile agent aims to reach a privately selected…

系统与控制 · 电气工程与系统科学 2026-05-19 Violetta Rostobaya , Yue Guan , James Berneburg , Daigo Shishika

We show how the software Walnut can be used to obtain concise proofs of results concerning variants of the famous Wythoff game, in which blocking maneuvers or terminal positions are added, as discussed respectively by Larsson (2011) and…

离散数学 · 计算机科学 2025-12-15 Antoine Renard , Michel Rigo

Optimising queries in real-world situations under imperfect conditions is still a problem that has not been fully solved. We consider finding the optimal order in which to execute a given set of selection operators under partial ignorance…

数据库 · 计算机科学 2015-07-30 Khaled H. Alyoubi , Sven Helmer , Peter T. Wood

We consider a setting in which a principal gets to choose which game from some given set is played by a group of agents. The principal would like to choose a game that favors one of the players, the social preferences of the players, or the…

计算机科学与博弈论 · 计算机科学 2025-11-27 Caspar Oesterheld , Vincent Conitzer

In a social system, the self-interest of agents can be detrimental to the collective good, sometimes leading to social dilemmas. To resolve such a conflict, a central designer may intervene by either redesigning the system or incentivizing…

机器学习 · 计算机科学 2020-10-28 Jiayang Li , Jing Yu , Yu Marco Nie , Zhaoran Wang

Infinitely repeated games support equilibrium concepts beyond those present in one-shot games (e.g., cooperation in the prisoner's dilemma). Nonetheless, repeated games fail to capture our real-world intuition for settings with many…

计算机科学与博弈论 · 计算机科学 2025-05-22 Henry Fleischmann , Kiriaki Fragkia , Ratip Emin Berker

We study rational agents with different perception capabilities in strategic games. We focus on a class of one-shot limited-perception games. These games extend simultaneous-move normal-form games by presenting each player with an…

计算机科学与博弈论 · 计算机科学 2024-05-28 Kai Jia , Martin Rinard

In reinforcement learning, temporal abstraction in the action space, exemplified by action repetition, is a technique to facilitate policy learning through extended actions. However, a primary limitation in previous studies of action…

机器学习 · 计算机科学 2024-02-09 Joongkyu Lee , Seung Joon Park , Yunhao Tang , Min-hwan Oh

Ensuring robust decision-making in multi-agent systems is challenging when agents have distinct, possibly conflicting objectives and lack full knowledge of each other's strategies. This is apparent in safety-critical applications such as…

系统与控制 · 电气工程与系统科学 2025-10-20 Francesco Bianchin , Robert Lefringhausen , Elisa Gaetan , Samuel Tesfazgi , Sandra Hirche

Decision making in modern large-scale and complex systems such as communication networks, smart electricity grids, and cyber-physical systems motivate novel game-theoretic approaches. This paper investigates big strategic (non-cooperative)…

计算机科学与博弈论 · 计算机科学 2016-09-22 Tansu Alpcan , Benjamin I. P. Rubinstein , Christopher Leckie

Landsburg method of classifying mixed Nash equilibria for maximally entangled Eisert-Lewenstein-Wilkens (ELW) game is analyzed with special emphasis on symmetries inherent to the problem. Nash equilibria for the original ELW game are…

量子物理 · 物理学 2016-03-28 Katarzyna Bolonek-Lasoń , Piotr Kosiński

Large Language Models (LLMs) like GPT-4 have revolutionized natural language processing, showing remarkable linguistic proficiency and reasoning capabilities. However, their application in strategic multi-agent decision-making environments…

计算与语言 · 计算机科学 2024-05-29 Chuanhao Li , Runhan Yang , Tiankai Li , Milad Bafarassat , Kourosh Sharifi , Dirk Bergemann , Zhuoran Yang

This paper examines games with strategic complements or substitutes and incomplete information, where players are uncertain about the opponents' parameters. We assume that the players' beliefs about the opponent's parameters are selected…

理论经济学 · 经济学 2025-01-28 Joep van Sloun

While many theoretical studies have revealed the strategies that could lead to and maintain cooperation in the Iterated Prisoner's Dilemma, less is known about what human participants actually do in this game and how strategies change when…

计算机科学与博弈论 · 计算机科学 2024-03-12 Eladio Montero-Porras , Jelena Grujic , Elias Fernandez-Domingos , Tom Lenaerts

Autonomous systems are increasingly expected to operate in the presence of adversaries, though adversaries may infer sensitive information simply by observing a system. Therefore, present a deceptive sequential decision-making framework…

Deep reinforcement learning agents are prone to goal misalignments. The black-box nature of their policies hinders the detection and correction of such misalignments, and the trust necessary for real-world deployment. So far, solutions…

人工智能 · 计算机科学 2024-05-27 Hector Kohler , Quentin Delfosse , Riad Akrour , Kristian Kersting , Philippe Preux