English
Related papers

Related papers: A conversion between utility and information

200 papers

Currently, the increase in financial returns from economic operations is constrained in view of the lack of a single efficiency criterion, which allows uniquely identify the business operation by their main feature - the possibility of…

Optimization and Control · Mathematics 2015-10-15 Igor Lutsenko

Energy considerations can significantly affect the behavior of a population of energy-consuming agents with limited energy budgets, for instance, in the movement process of people in a city. We consider a population of interacting agents…

Disordered Systems and Neural Networks · Physics 2024-11-13 Mohsen Ghasemi Nezhadhaghighi , Abolfazl Ramezanpour

In some agent designs like inverse reinforcement learning an agent needs to learn its own reward function. Learning the reward function and optimising for it are typically two different processes, usually performed at different stages. We…

Artificial Intelligence · Computer Science 2020-04-29 Stuart Armstrong , Jan Leike , Laurent Orseau , Shane Legg

In many real-world tasks, it is not possible to procedurally specify an RL agent's reward function. In such cases, a reward function must instead be learned from interacting with and observing humans. However, current techniques for reward…

Machine Learning · Computer Science 2020-12-11 Eric J. Michaud , Adam Gleave , Stuart Russell

We develop a tractable model of realization utility that studies the role of reference-dependent S-shaped preferences in a dynamic investment setting with reinvestment. Our model generates both voluntarily realized gains and losses. It…

General Finance · Quantitative Finance 2014-08-14 Jonathan E. Ingersoll , Lawrence J. Jin

In many real-world scenarios, rewards extrinsic to the agent are extremely sparse, or absent altogether. In such cases, curiosity can serve as an intrinsic reward signal to enable the agent to explore its environment and learn skills that…

Machine Learning · Computer Science 2017-05-16 Deepak Pathak , Pulkit Agrawal , Alexei A. Efros , Trevor Darrell

We consider the design of experiments to evaluate treatments that are administered by self-interested agents, each seeking to achieve the highest evaluation and win the experiment. For example, in an advertising experiment, a company wishes…

Methodology · Statistics 2015-09-18 Panos Toulis , David C. Parkes , Elery Pfeffer , James Zou

We introduce a utility-driven bounded-confidence model of opinion dynamics in which opinions associated with higher utility exert stronger social influence. In the regime where all agents belong to a single opinion cluster, we derive a…

Adaptation and Self-Organizing Systems · Physics 2026-05-22 Alex Siebenmorgen , Juan G. Restrepo

In order to compute near-optimal policies with policy-gradient algorithms, it is common in practice to include intrinsic exploration terms in the learning objective. Although the effectiveness of these terms is usually justified by an…

Machine Learning · Computer Science 2025-08-21 Adrien Bolland , Gaspard Lambrechts , Damien Ernst

We obtain an elementary characterization of expected utility based on a representation of choice in terms of psychological gambles, which requires no assumption other than coherence between ex-ante and ex-post preferences. Weaker version of…

General Economics · Economics 2024-11-05 Gianluca Cassese

In this paper, we propose an optimization-based mechanism to explain power law distributions, where the function that the optimization process is seeking to optimize is derived mathematically, then the behavior and interpretation of this…

Physics and Society · Physics 2018-12-27 A. M. Khalili

[Context and motivation:] For realistic self-adaptive systems, multiple quality attributes need to be considered and traded off against each other. These quality attributes are commonly encoded in a utility function, for instance, a…

Software Engineering · Computer Science 2021-03-19 Rebekka Wohlrab , David Garlan

We consider an analyst whose goal is to identify a subject's utility function through revealed preference analysis. We argue the analyst's preference about which experiments to run should adhere to three normative principles: The first,…

Theoretical Economics · Economics 2024-11-19 Fernando Payró , Evan Piermont

This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quantities are decision-theoretic, but have information-theoretic analogues, e.g., they share some…

Information Theory · Computer Science 2026-01-13 Aolin Xu

In reinforcement learning (RL), agents sequentially interact with changing environments while aiming to maximize the obtained rewards. Usually, rewards are observed only after acting, and so the goal is to maximize the expected cumulative…

Machine Learning · Computer Science 2024-10-15 Nadav Merlis , Dorian Baudry , Vianney Perchet

We study online learning for optimal allocation when the resource to be allocated is time. %Examples of possible applications include job scheduling for a computing server, a driver filling a day with rides, a landlord renting an estate,…

Machine Learning · Statistics 2021-11-05 Etienne Boursier , Tristan Garrec , Vianney Perchet , Marco Scarsini

We provide sufficient conditions under which a utility function may be recovered from a finite choice experiment. Identification, as is commonly understood in decision theory, is not enough. We provide a general recoverability result that…

Theoretical Economics · Economics 2023-01-30 Christopher P. Chambers , Federico Echenique , Nicolas S. Lambert

The incentive ratio measures the utility gains from strategic behaviour. Without any restrictions on the setup, ratios for linear, Leontief and Cobb-Douglas exchange markets are unbounded, showing that manipulating the equilibrium is a…

Computer Science and Game Theory · Computer Science 2017-05-01 Ido Polak

We study how individuals trade off outcome ("what") and process ("how") utility in high-stakes strategic decisions, namely professional tennis. Using optimality conditions and the second-service rule, we derive a sufficient condition for…

Econometrics · Economics 2026-05-25 Arnaud Dupuy

When deploying autonomous agents in the real world, we need effective ways of communicating objectives to them. Traditional skill learning has revolved around reinforcement and imitation learning, each with rigid constraints on the format…

Artificial Intelligence · Computer Science 2019-11-21 Mark Woodward , Chelsea Finn , Karol Hausman