English
Related papers

Related papers: Policy Based Inference in Trick-Taking Card Games

200 papers

In the multiarmed bandit problem a gambler chooses an arm of a slot machine to pull considering a tradeoff between exploration and exploitation. We study the stochastic bandit problem where each arm has a reward distribution supported in a…

Statistics Theory · Mathematics 2013-03-29 Junya Honda , Akimichi Takemura

Recent advances in reinforcement learning with social agents have allowed such models to achieve human-level performance on specific interaction tasks. However, most interactive scenarios do not have a version alone as an end goal; instead,…

Artificial Intelligence · Computer Science 2022-08-23 Pablo Barros , Ozge Nilay Yalcın , Ana Tanevska , Alessandra Sciutti

We introduce a new virtual environment for simulating a card game known as "Big 2". This is a four-player game of imperfect information with a relatively complicated action space (being allowed to play 1,2,3,4 or 5 card combinations from an…

Machine Learning · Computer Science 2018-09-03 Henry Charlesworth

Finding a best response policy is a central objective in game theory and multi-agent learning, with modern population-based training approaches employing reinforcement learning algorithms as best-response oracles to improve play against…

Artificial Intelligence · Computer Science 2023-04-26 Daniel Hernandez , Hendrik Baier , Michael Kaisers

When learning to play an imperfect information game, it is often easier to first start with the basic mechanics of the game rules. For example, one can play several example rounds with private cards revealed to all players to better…

Computer Science and Game Theory · Computer Science 2025-05-27 Benjamin Heymann , Marc Lanctot

It is suggested that an AI inference system should reflect an inference policy that is tailored to the domain of problems to which it is applied -- and furthermore that an inference policy need not conform to any general theory of rational…

Artificial Intelligence · Computer Science 2013-04-08 Paul E. Lehner

Game AI systems need the theory of mind, which is the humanistic ability to infer others' mental models, preferences, and intent. Such systems would enable inferring players' behavior tendencies that contribute to the variations in their…

Human-Computer Interaction · Computer Science 2021-07-27 Murtuza N. Shergadwala , Zhaoqing Teng , Magy Seif El-Nasr

A general graph-structured neural network architecture operates on graphs through two core components: (1) complex enough message functions; (2) a fixed information aggregation process. In this paper, we present the Policy Message Passing…

Machine Learning · Computer Science 2019-10-01 Zhiwei Deng , Greg Mori

We propose a generic mechanism for incentivizing behavior in an arbitrary finite game using payments. Doing so is trivial if the mechanism is allowed to observe all actions taken in the game, as this allows it to simply punish those agents…

Computer Science and Game Theory · Computer Science 2023-04-05 Nikolaj I. Schwartzbach

Policy inference plays an essential role in the contextual bandit problem. In this paper, we use empirical likelihood to develop a Bayesian inference method for the joint analysis of multiple contextual bandit policies in finite sample…

Machine Learning · Statistics 2026-02-12 Jiangrong Ouyang , Mingming Gong , Howard Bondell

Humans can learn many novel tasks from a very small number (1--5) of demonstrations, in stark contrast to the data requirements of nearly tabula rasa deep learning methods. We propose an expressive class of policies, a strong but general…

Artificial Intelligence · Computer Science 2019-11-19 Tom Silver , Kelsey R. Allen , Alex K. Lew , Leslie Pack Kaelbling , Josh Tenenbaum

This paper explores successor features for knowledge transfer in zero-sum, complete-information, and turn-based games. Prior research in single-agent systems has shown that successor features can provide a ``jump start" for agents when…

Multiagent Systems · Computer Science 2025-07-31 Sunny Amatya , Yi Ren , Zhe Xu , Wenlong Zhang

Understanding, predicting, and learning from other people's actions are fundamental human social-cognitive skills. Little is known about how and when we consider other's actions and outcomes when making our own decisions. We developed a…

Neurons and Cognition · Quantitative Biology 2019-05-21 Julia Anna Adrian , Siddharth Siddharth , Syed Zain Ali Baquar , Tzyy-Ping Jung , Gedeon Deák

Deceptive patterns are often used in interface design to manipulate users into taking actions they would not otherwise take, such as consenting to excessive data collection. We present Trickery, a narrative serious game that incorporates…

Human-Computer Interaction · Computer Science 2024-10-25 Kirill Kronhardt , Kevin Rolfes , Jens Gerken

Power indices are essential in assessing the contribution and influence of individual agents in multi-agent systems, providing crucial insights into collaborative dynamics and decision-making processes. While invaluable, traditional…

Multiagent Systems · Computer Science 2025-03-12 Benjamin Kempinski , Tal Kachman

We propose the first boosting algorithm for off-policy learning from logged bandit feedback. Unlike existing boosting methods for supervised learning, our algorithm directly optimizes an estimate of the policy's expected reward. We analyze…

Machine Learning · Computer Science 2023-05-03 Ben London , Levi Lu , Ted Sandler , Thorsten Joachims

While game-theoretic planning frameworks are effective at modeling multi-agent interactions, they require solving large optimization problems where the number of variables increases with the number of agents, resulting in long computation…

Robotics · Computer Science 2026-04-02 Tianyu Qiu , Eric Ouano , Fernando Palafox , Christian Ellis , David Fridovich-Keil

This study employs gamified experiments to investigate and refine the Schelling Model of Segregation, a framework that demonstrates how individual preferences can lead to systemic segregation. Using a movement selection algorithm derived…

Physics and Society · Physics 2025-01-15 Aleix Nicolás Olivé , Luce Prignano , Dimitri Marinelli , Emanuele Cozzo

We define a new concept of "mistake" strategies and actions for strategic-form and extensive-form games, analyze the relationship to prior main game-theoretic solution concepts, study algorithms for computation, and explore practicality.…

Computer Science and Game Theory · Computer Science 2020-10-29 Sam Ganzfried

Predicting player behavior in strategic games, especially complex ones like chess, presents a significant challenge. The difficulty arises from several factors. First, the sheer number of potential outcomes stemming from even a single…

Machine Learning · Computer Science 2025-04-09 Benny Skidanov , Daniel Erbesfeld , Gera Weiss , Achiya Elyasaf