English
Related papers

Related papers: Adaptive Multi-Head Finite-State Gamblers

200 papers

We demonstrate the use of a multidimensional extension of the latent Markov model to analyse data from studies with correlated binary responses in developmental psychology. In particular, we consider an experiment based on a battery of…

Applications · Statistics 2025-01-08 Francesco Bartolucci , Ivonne L. Solis-Trapala

We consider adaptive decision-making problems where an agent optimizes a cumulative performance objective by repeatedly choosing among a finite set of options. Compared to the classical prediction-with-expert-advice set-up, we consider…

Machine Learning · Computer Science 2023-04-10 Michael Muehlebach

We develop an approach to solve Barberis (2012)'s casino gambling model in which a gambler whose preferences are specified by the cumulative prospect theory (CPT) must decide when to stop gambling by a prescribed deadline. We assume that…

Mathematical Finance · Quantitative Finance 2021-02-08 Sang Hu , Jan Obloj , Xun Yu Zhou

Observational learning often involves congestion: an agent gets lower payoff from an action when more predecessors have taken that action. This preference to act differently from previous agents may paradoxically increase all but one…

Theoretical Economics · Economics 2019-04-02 Sander Heinsalu

Empirical evidence shows that human behaviour often deviates from game-theoretical rationality. For instance, humans may hold unrealistic expectations about future outcomes. As the evolutionary roots of such biases remain unclear, we…

Multiagent Systems · Computer Science 2025-08-29 Marco Saponara , Elias Fernandez Domingos , Jorge M. Pacheco , Tom Lenaerts

We are interested in the convergence of the value of n-stage games as n goes to infinity and the existence of the uniform value in stochastic games with a general set of states and finite sets of actions where the transition is commutative.…

Optimization and Control · Mathematics 2016-04-22 Xavier Venel

Multivariate Hawkes processes are past-dependant point processes originally introduced to model excitation effects, later extended to a nonlinear framework to account for the opposite effect, known as inhibition. Motivated by applications…

Methodology · Statistics 2026-05-12 Sacha Quayle , Anna Bonnet , Maxime Sangnier

Despite substantial progress in recent years, probabilistic solvers with adaptive step sizes can still not solve memory-demanding differential equations -- unless we care only about a single point in time (which is far too restrictive; we…

Numerical Analysis · Mathematics 2025-07-04 Nicholas Krämer

We show that under some general conditions the finite memory determinacy of a class of two-player win/lose games played on finite graphs implies the existence of a Nash equilibrium built from finite memory strategies for the corresponding…

Computer Science and Game Theory · Computer Science 2016-07-13 Stéphane Le Roux , Arno Pauly

Online learning algorithms that minimize regret provide strong guarantees in situations that involve repeatedly making decisions in an uncertain environment, e.g. a driver deciding what route to drive to work every day. While regret…

Computer Science and Game Theory · Computer Science 2013-09-06 Jeremiah Blocki , Nicolas Christin , Anupam Datta , Arunesh Sinha

This paper investigates the effect of learning a forward model on the performance of a statistical forward planning agent. We transform Conway's Game of Life simulation into a single-player game where the objective can be either to preserve…

In decision-dependent games, multiple players optimize their decisions under a data distribution that shifts with their joint actions, creating complex dynamics in applications like market pricing. A practical consequence of these dynamics…

Computer Science and Game Theory · Computer Science 2025-09-04 Guangzheng Zhong , Yang Liu , Jiming Liu

We formulate an adaptive version of Kelly's horse model in which the gambler learns from past race results using Bayesian inference. A known asymptotic scaling for the difference between the growth rate of the gambler and the optimal growth…

Statistical Mechanics · Physics 2022-10-05 Armand Despons , David Lacoste , Luca Peliti

Multi-hop inference is necessary for machine learning systems to successfully solve tasks such as Recognising Textual Entailment and Machine Reading. In this work, we demonstrate the effectiveness of adaptive computation for learning the…

Computation and Language · Computer Science 2016-11-17 Mark Neumann , Pontus Stenetorp , Sebastian Riedel

Predicting outcomes in sports is important for teams, leagues, bettors, media, and fans. Given the growing amount of player tracking data, sports analytics models are increasingly utilizing spatially-derived features built upon player…

Machine Learning · Computer Science 2022-07-29 Peter Xenopoulos , Claudio Silva

The two-dimensional Hubbard model at finite doping hosts competing or intertwined orders, resulting in conflicting conclusions from different computational approaches regarding its ground state. We show that a key source of such…

Strongly Correlated Electrons · Physics 2026-04-27 Luciano Loris Viteritti , Riccardo Rende , Christopher Roth , Anirvan Sengupta , Giuseppe Carleo , Antoine Georges

A puzzle about prisoners trying to identify the color of a hat on their head leads to a version where there are k more hats than prisoners. This generalized puzzle is related to the independence number of the arrangement graph A(m, n) and…

Combinatorics · Mathematics 2019-03-25 Rob Pratt , Stan Wagon , Michael Wiener , Piotr Zielinski

Defensive coverage schemes in the National Football League (NFL) represent complex tactical patterns requiring coordinated assignments among defenders who must react dynamically to the offense's passing concept. This paper presents a…

Machine Learning · Computer Science 2026-03-30 Kevin Song , Evan Diewald , Ornob Siddiquee , Chris Boomhower , Keegan Abdoo , Mike Band , Amy Lee

Learning models do not in general imply that weakly dominated strategies are irrelevant or justify the related concept of "forward induction," because rational agents may use dominated strategies as experiments to learn how opponents play,…

Theoretical Economics · Economics 2022-11-15 Daniel Clark , Drew Fudenberg , Kevin He

Predicting high-fidelity future human poses, from a historically observed sequence, is decisive for intelligent robots to interact with humans. Deep end-to-end learning approaches, which typically train a generic pre-trained model on…

Computer Vision and Pattern Recognition · Computer Science 2023-04-14 Qiongjie Cui , Huaijiang Sun , Jianfeng Lu , Bin Li , Weiqing Li