English
Related papers

Related papers: Convergence of Expected Utility for Universal AI

200 papers

This paper investigates the problem of maximizing expected terminal utility in a discrete-time financial market model with a finite horizon under non-dominated model uncertainty. We use a dynamic programming framework together with…

Mathematical Finance · Quantitative Finance 2017-10-03 Laurence Carassus , Romain Blanchard

This paper concerns the estimation of sums of functions of observable and unobservable variables. Lower bounds for the asymptotic variance and a convolution theorem are derived in general finite- and infinite-dimensional models. An explicit…

Statistics Theory · Mathematics 2007-06-13 Cun-Hui Zhang

An analyst observes an agent take a sequence of actions. The analyst does not have access to the agent's information and ponders whether the observed actions could be justified through a rational Bayesian model with a known utility…

Theoretical Economics · Economics 2025-04-08 Henrique de Oliveira , Rohit Lamba

The analysis of the adaptive behaviour of many different kinds of systems such as humans, animals and machines, requires more general ways of assessing their cognitive abilities. This need is strengthened by increasingly more tasks being…

Artificial Intelligence · Computer Science 2013-05-10 David L. Dowe , Jose Hernandez-Orallo

Game-theoretic dynamics between AI agents could differ from traditional human-human interactions in various ways. One such difference is that it may be possible to accurately simulate an AI agent, for example because its source code is…

Artificial Intelligence · Computer Science 2024-03-05 Vojtech Kovarik , Caspar Oesterheld , Vincent Conitzer

Is transparency always beneficial in complex systems such as traffic networks and stock markets? How is transparency defined in multi-agent systems, and what is its optimal degree at which social welfare is highest? We take an agent-based…

Multiagent Systems · Computer Science 2024-01-12 Kshama Dwarakanath , Svitlana Vyetrenko , Toks Oyebode , Tucker Balch

We apply the maximum entropy principle to economic systems in equilibrium and find the density function for the market's wealth. This is the same as price density which is used for insurance pricing. The risk aversion parameter of the agent…

Statistical Mechanics · Physics 2008-12-10 Amir H. Darooneh

We present several new characterizations of correlated equilibria in games with continuous utility functions. These have the advantage of being more computationally and analytically tractable than the standard definition in terms of…

Computer Science and Game Theory · Computer Science 2011-06-06 Noah D. Stein , Pablo A. Parrilo , Asuman Ozdaglar

In the Bayesian approach to sequential decision making, exact calculation of the (subjective) utility is intractable. This extends to most special cases of interest, such as reinforcement learning problems. While utility bounds are known to…

Machine Learning · Computer Science 2011-11-14 Christos Dimitrakakis

We provide and axiomatize a representation for preferences over lotteries that generalizes the expected utility model. Since the representation uses different utility functions to evaluate different lotteries, the preferences can be…

Theoretical Economics · Economics 2026-03-17 Edward Honda , Keh-Kuan Sun

Policy learning is an important component of many real-world learning systems. A major challenge in policy learning is how to adapt efficiently to unseen environments or tasks. Recently, it has been suggested to exploit invariant…

Machine Learning · Statistics 2023-06-28 Sorawit Saengkyongam , Niklas Pfister , Predrag Klasnja , Susan Murphy , Jonas Peters

Predicting the outcomes of cyber-physical systems with multiple human interactions is a challenging problem. This article reviews a game theoretical approach to address this issue, where reinforcement learning is employed to predict the…

Multiagent Systems · Computer Science 2019-10-14 Mert Albaba , Yildiray Yildiz

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

Machine Learning · Computer Science 2014-08-12 Aristide Tossou , Christos Dimitrakakis

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

Machine Learning · Statistics 2013-07-16 Aristide C. Y. Tossou , Christos Dimitrakakis

Adaptive machines have the potential to assist or interfere with human behavior in a range of contexts, from cognitive decision-making to physical device assistance. Therefore it is critical to understand how machine learning algorithms can…

Artificial Intelligence · Computer Science 2023-05-03 Benjamin J. Chasnov , Lillian J. Ratliff , Samuel A. Burden

We explore a collaborative and cooperative multi-agent reinforcement learning setting where a team of reinforcement learning agents attempt to solve a single cooperative task in a multi-scenario setting. We propose a novel multi-agent…

Multiagent Systems · Computer Science 2019-08-27 Hassam Ullah Sheikh , Ladislau Bölöni

In cases of uncertainty, a multi-class classifier preferably returns a set of candidate classes instead of predicting a single class label with little guarantee. More precisely, the classifier should strive for an optimal balance between…

Machine Learning · Computer Science 2020-05-28 Thomas Mortier , Marek Wydmuch , Krzysztof Dembczyński , Eyke Hüllermeier , Willem Waegeman

The predictability of a sequence is defined as the asymptotic performance of the best performing predictor in a given class. The value of the predictability of a sequence will in general depend on the choice of this predictor class. The…

Statistics Theory · Mathematics 2009-04-15 Finn Macleod , Alexei Pokrovskii , Dima Rachinskii

In dynamic settings each economic agent's choices can be revealing of her private information. This elicitation via the rationalization of observable behavior depends each agent's perception of which payoff-relevant contingencies other…

Theoretical Economics · Economics 2021-05-17 Evan Piermont , Peio Zuazo-Garin

For safe and reliable deployment in the real world, autonomous agents must elicit appropriate levels of trust from human users. One method to build trust is to have agents assess and communicate their own competencies for performing given…

Robotics · Computer Science 2022-06-22 Aastha Acharya , Rebecca Russell , Nisar R. Ahmed
‹ Prev 1 8 9 10 Next ›