English
Related papers

Related papers: Policy Learning with $\alpha$-Expected Welfare

200 papers

This paper examines optimal risk sharing for empirically realistic risk attitudes, providing results on Pareto optimality, competitive equilibria, utility frontiers, and the first and second theorems of welfare. Contrary to common…

Theoretical Economics · Economics 2025-10-06 Jean-Gabriel Lauzier , Liyuan Lin , Peter Wakker , Ruodu Wang

We consider off-policy policy evaluation with function approximation (FA) in average-reward MDPs, where the goal is to estimate both the reward rate and the differential value function. For this problem, bootstrapping is necessary and,…

Machine Learning · Computer Science 2022-10-19 Shangtong Zhang , Yi Wan , Richard S. Sutton , Shimon Whiteson

Policy learning can be used to extract individualized treatment regimes from observational data in healthcare, civics, e-commerce, and beyond. One big hurdle to policy learning is a commonplace lack of overlap in the data for different…

Machine Learning · Statistics 2020-12-04 Nathan Kallus

In this paper, we propose a general operating scheme which allows the utility company to jointly perform power procurement and demand response so as to maximize the social welfare. Our model takes into consideration the effect of the…

Optimization and Control · Mathematics 2011-12-06 Longbo Huang , Jean Walrand , Kannan Ramchandran

We provide theoretical investigations into off-policy evaluation in reinforcement learning using function approximators for (marginalized) importance weights and value functions. Our contributions include: (1) A new estimator, MWL, that…

Machine Learning · Computer Science 2020-10-08 Masatoshi Uehara , Jiawei Huang , Nan Jiang

Statistical parity metrics have been widely studied and endorsed in the AI community as a means of achieving fairness, but they suffer from at least two weaknesses. They disregard the actual welfare consequences of decisions and may…

Artificial Intelligence · Computer Science 2024-05-21 Violet Chen , J. N. Hooker , Derek Leben

We investigate the problem of best policy identification in discounted linear Markov Decision Processes in the fixed confidence setting under a generative model. We first derive an instance-specific lower bound on the expected number of…

Machine Learning · Computer Science 2022-08-12 Jerome Taupin , Yassir Jedra , Alexandre Proutiere

In this paper we develop an Expectation Maximization(EM) algorithm to estimate the parameter of a Yule-Simon distribution. The Yule-Simon distribution exhibits the "rich get richer" effect whereby an 80-20 type of rule tends to dominate.…

Computation · Statistics 2020-11-17 Lucas Roberts , Denisa Roberts

Machine learning is increasingly used in government programs to identify and support the most vulnerable individuals, prioritizing assistance for those at greatest risk over optimizing aggregate outcomes. This paper examines the welfare…

Computers and Society · Computer Science 2025-07-14 Unai Fischer-Abaigar , Christoph Kern , Juan Carlos Perdomo

We study the fair allocation of indivisible items to $n$ agents to maximize the utilitarian social welfare, where the fairness criterion is envy-free up to one item and there are only two different utility functions shared by the agents. We…

Computer Science and Game Theory · Computer Science 2025-09-12 Jiaxuan Ma , Yong Chen , Guangting Chen , Mingyang Gong , Guohui Lin , An Zhang

The EM (Expectation-Maximization) algorithm is regarded as an MM (Majorization-Minimization) algorithm for maximum likelihood estimation of statistical models. Expanding this view, this paper demonstrates that by choosing an appropriate…

Optimization and Control · Mathematics 2026-02-12 Kensuke Asai , Jun-ya Gotoh

A policymaker discloses public information to interacting agents who also acquire costly private information. More precise public information reduces the precision and cost of acquired private information. Considering this effect, what…

Theoretical Economics · Economics 2022-04-08 Takashi Ui

We consider the problem of estimating personalized treatment policies that are "externally valid" or "generalizable": they perform well in target populations that differ from the experimental (or training) population from which the data are…

Econometrics · Economics 2025-11-10 Christopher Adjaho , Timothy Christensen

For the fundamental problem of allocating a set of resources among individuals with varied preferences, the quality of an allocation relates to the degree of fairness and the collective welfare achieved. Unfortunately, in many…

Computer Science and Game Theory · Computer Science 2024-08-30 Mikael Møller Høgsgaard , Panagiotis Karras , Wenyue Ma , Nidhi Rathi , Chris Schwiegelshohn

Strong empirical evidence from laboratory experiments, and more recently from population surveys, shows that individuals, when evaluating their situations, pay attention to whether they experience gains or losses, with losses weighing more…

Theoretical Economics · Economics 2025-10-17 Martyna Kobus , Radosław Kurek , Thomas Parker

A set of divisible resources becomes available over a sequence of rounds and needs to be allocated immediately and irrevocably. Our goal is to distribute these resources to maximize fairness and efficiency. Achieving any non-trivial…

Computer Science and Game Theory · Computer Science 2020-09-29 Vasilis Gkatzelis , Alexandros Psomas , Xizhi Tan

In the pursuit of finding an optimal policy, reinforcement learning (RL) methods generally ignore the properties of learned policies apart from their expected return. Thus, even when successful, it is difficult to characterize which…

Machine Learning · Computer Science 2025-10-10 Yash Jhaveri , Harley Wiltzer , Patrick Shafto , Marc G. Bellemare , David Meger

The assumption of normality in data has been considered in the field of statistical analysis for a long time. However, in many practical situations, this assumption is clearly unrealistic. It has recently been suggested that the use of…

Computation · Statistics 2016-11-25 Reinaldo B. Arellano-Valle , Javier E. Contreras-Reyes

When a policy prioritizes one person over another, is it because they benefit more, or because they are preferred? This paper develops a method to uncover the values consistent with observed allocation decisions. We use machine learning…

General Economics · Economics 2022-06-03 Daniel Björkegren , Joshua E. Blumenstock , Samsun Knight

I consider a class of statistical decision problems in which the policymaker must decide between two policies to maximize social welfare (e.g., the population mean of an outcome) based on a finite sample. The framework introduced in this…

Econometrics · Economics 2025-03-04 Kohei Yata
‹ Prev 1 4 5 6 7 8 10 Next ›