English
Related papers

Related papers: A conversion between utility and information

200 papers

Energy prices and net power injection limitations regulate the operations in distribution grids and typically ensure that operational constraints are met. Nevertheless, unexpected or prolonged abnormal events could undermine the grid's…

Systems and Control · Electrical Eng. & Systems 2024-03-29 Guido Cavraro , Joshua Comden , Andrey Bernstein

It is often difficult to hand-specify what the correct reward function is for a task, so researchers have instead aimed to learn reward functions from human behavior or feedback. The types of behavior interpreted as evidence of the reward…

Machine Learning · Computer Science 2020-12-14 Hong Jun Jeon , Smitha Milli , Anca D. Dragan

This paper develops generalizations of empowerment to continuous states. Empowerment is a recently introduced information-theoretic quantity motivated by hypotheses about the efficiency of the sensorimotor loop in biological organisms, but…

Artificial Intelligence · Computer Science 2012-02-01 Tobias Jung , Daniel Polani , Peter Stone

The modernization of the power system introduces technologies that may improve the system's efficiency by enhancing the capabilities of users. Despite their potential benefits, such technologies can have a negative impact. This subject has…

Computer Science and Game Theory · Computer Science 2019-06-28 Carlos Barreto , Eduardo Mojica-Nava , Nicanor Quijano

Efficient exploration remains a challenging problem in reinforcement learning, especially for those tasks where rewards from environments are sparse. A commonly used approach for exploring such environments is to introduce some "intrinsic"…

Machine Learning · Computer Science 2020-07-16 Neale Ratzlaff , Qinxun Bai , Li Fuxin , Wei Xu

Preference-based reward learning is a popular technique for teaching robots and autonomous systems how a human user wants them to perform a task. Previous works have shown that actively synthesizing preference queries to maximize…

Robotics · Computer Science 2024-03-12 Evan Ellis , Gaurav R. Ghosal , Stuart J. Russell , Anca Dragan , Erdem Bıyık

Reward schemes may affect not only agents' effort, but also their incentives to gather information to reduce the riskiness of the productive activity. In a laboratory experiment using a novel task, we find that the relationship between…

General Economics · Economics 2024-09-11 Philip Brookins , Jennifer Brown , Dmitry Ryvkin

We discuss representing and reasoning with knowledge about the time-dependent utility of an agent's actions. Time-dependent utility plays a crucial role in the interaction between computation and action under bounded resources. We present a…

Artificial Intelligence · Computer Science 2013-03-26 Eric J. Horvitz , Geoffrey Rutledge

The inputs and preferences of human users are important considerations in situations where these users interact with autonomous cyber or cyber-physical systems. In these scenarios, one is often interested in aligning behaviors of the system…

Machine Learning · Computer Science 2021-04-02 Bhaskar Ramasubramanian , Luyao Niu , Andrew Clark , Radha Poovendran

This paper studies the optimal mechanism to motivate effort in a dynamic principal-agent model without transfers. An agent is engaged in a task with uncertain future rewards and can quit at any time. The principal knows the reward and…

Theoretical Economics · Economics 2026-01-16 Chang Liu

Since the 1960s, the question whether markets are efficient or not is controversially discussed. One reason for the difficulty to overcome the controversy is the lack of a universal, but also precise, quantitative definition of efficiency…

General Finance · Quantitative Finance 2018-12-10 Roland Rothenstein

Shannon's information entropy measures of the uncertainty of an event's outcome. If learning about a system reflects a decrease in uncertainty, then a plausible intuition is that learning should be accompanied by a decrease in the entropy…

Robotics · Computer Science 2015-02-20 Paul E. Smaldino

We investigate how effective an attacker can be when it only learns from its victim's actions, without access to the victim's reward. In this work, we are motivated by the scenario where the attacker wants to behave strategically when the…

Machine Learning · Computer Science 2021-12-03 Ted Fujimoto , Timothy Doster , Adam Attarian , Jill Brandenberger , Nathan Hodas

In dynamic settings each economic agent's choices can be revealing of her private information. This elicitation via the rationalization of observable behavior depends each agent's perception of which payoff-relevant contingencies other…

Theoretical Economics · Economics 2021-05-17 Evan Piermont , Peio Zuazo-Garin

Active inference proposes expected free energy as an objective for planning and decision-making to adequately balance exploitative and explorative drives in learning agents. The exploitative drive, or what an agent wants to achieve, is…

Artificial Intelligence · Computer Science 2025-12-04 Filippo Torresan , Ryota Kanai , Manuel Baltieri

Typically, merit is defined with respect to some intrinsic measure of worth. We instead consider a setting where an individual's worth is \emph{relative}: when a Decision Maker (DM) selects a set of individuals from a population to maximise…

Artificial Intelligence · Computer Science 2022-09-12 Thomas Kleine Buening , Meirav Segal , Debabrota Basu , Christos Dimitrakakis , Anne-Marie George

Instead of testing for unanimous agreement, I propose learning how broad of a consensus favors one distribution over another (of earnings, productivity, asset returns, test scores, etc.). Specifically, given a sample from each of two…

Econometrics · Economics 2024-08-27 David M. Kaplan

Explainable AI techniques that describe agent reward functions can enhance human-robot collaboration in a variety of settings. One context where human understanding of agent reward functions is particularly beneficial is in the value…

Robotics · Computer Science 2021-10-11 Lindsay Sanneman , Julie Shah

In barter exchanges agents enter seeking to swap their items for other items on their wishlist. We consider a centralized barter exchange with a set of agents and items where each item has a positive value. The goal is to compute a…

Data Structures and Algorithms · Computer Science 2024-06-21 Juan Luque , Sharmila Duppala , John Dickerson , Aravind Srinivasan

We consider multi-agent systems with general information networks where an agent may only observe a subset of other agents. A system designer assigns local utility functions to the agents guiding their actions towards an outcome which…

Computer Science and Game Theory · Computer Science 2025-01-30 Vartika Singh , Will Wesley , Philip N. Brown