English
Related papers

Related papers: Aggregating time preferences with decreasing impat…

200 papers

We study the problem of preferential Bayesian optimization (BO), where we aim to optimize a black-box function with only preference feedback over a pair of candidate solutions. Inspired by the likelihood ratio idea, we construct a…

Machine Learning · Computer Science 2024-05-30 Wenjie Xu , Wenbin Wang , Yuning Jiang , Bratislav Svetozarevic , Colin N. Jones

We consider a discrete-time version of the popular optimal dividend pay-out problem in risk theory. The novel aspect of our approach is that we allow for a risk averse insurer, i.e., instead of maximising the expected discounted dividends…

Probability · Mathematics 2015-12-02 Nicole Bäuerle , Anna Jaśkiewicz

It is shown that in the case of a single decision maker who optimizes several possibly conflicting objectives, the amount of information available in preference relations among pairs of possible decisions, when compared with all other…

Optimization and Control · Mathematics 2007-05-23 Elemer E Rosinger

Temporal difference (TD) learning is an important approach in reinforcement learning, as it combines ideas from dynamic programming and Monte Carlo methods in a way that allows for online and incremental model-free learning. A key idea of…

Machine Learning · Computer Science 2018-09-21 Kristopher De Asis , Brendan Bennett , Richard S. Sutton

Folklore suggests that policy gradient can be more robust to misspecification than its relative, approximate policy iteration. This paper studies the case of state-aggregated representations, where the state space is partitioned and either…

Machine Learning · Computer Science 2022-06-24 Daniel Russo

Hyperbolic decay time series such as, fractional Gaussian noise (FGN) or fractional autoregressive moving-average (FARMA) process, each exhibit two distinct types of behaviour: strong persistence or antipersistence. Beran (1994)…

Statistics Theory · Mathematics 2016-11-04 A. Ian McLeod

We build upon recent work [Kleinberg and Oren, 2014, Kleinberg et al., 2016, 2017] that considers present biased agents, who place more weight on costs they must incur now than costs they will incur in the future. They consider a graph…

Computer Science and Game Theory · Computer Science 2022-01-17 Aditya Saraf , Anna R. Karlin , Jamie Morgenstern

We explore questions dealing with the learnability of models of choice over time. We present a large class of preference models defined by a structural criterion for which we are able to obtain an exponential improvement over previously…

Computer Science and Game Theory · Computer Science 2018-09-11 Zachary Chase , Siddharth Prasad

We assess the demand effects of discounts on train tickets issued by the Swiss Federal Railways, the so-called `supersaver tickets', based on machine learning, a subfield of artificial intelligence. Considering a survey-based sample of…

General Economics · Economics 2022-07-01 Martin Huber , Jonas Meier , Hannes Wallimann

Interactive preference elicitation (IPE) aims to substantially reduce human effort while acquiring human preferences in wide personalization systems. Dueling bandit (DB) algorithms enable optimal decision-making in IPE building on pairwise…

Machine Learning · Computer Science 2025-11-13 Shengbo Wang , Hong Sun , Ke Li

This work addresses the output consensus problem of constrained heterogeneous multi-agent systems under a switching network with potential communication delays, where outputs are periodic and characterized by an exosystem. Since periodic…

Systems and Control · Electrical Eng. & Systems 2026-04-14 Shibo Han , Bonan Hou , Chong Jin Ong

Humans exhibit time-inconsistent behavior, in which planned actions diverge from executed actions. Understanding time inconsistency and designing appropriate interventions is a key research challenge in computer science and behavioral…

Computer Science and Game Theory · Computer Science 2025-09-18 Yasunori Akagi , Takeshi Kurashima

This paper provides a behavioral analysis of conservatism in beliefs. I introduce a new axiom, Dynamic Conservatism, that relaxes Dynamic Consistency when information and prior beliefs "conflict." When the agent is a subjective expected…

Theoretical Economics · Economics 2021-02-02 Matthew Kovach

Time-inconsistency is a characteristic of human behavior in which people plan for long-term benefits but take actions that differ from the plan due to conflicts with short-term benefits. Such time-inconsistent behavior is believed to be…

Computer Science and Game Theory · Computer Science 2025-01-15 Yasunori Akagi , Naoki Marumo , Takeshi Kurashima

We consider an economic agent (a household or an insurance company) modelling its surplus process by a deterministic process or by a Brownian motion with drift. The goal is to maximise the expected discounted spendings/dividend payments,…

Mathematical Finance · Quantitative Finance 2018-09-03 Julia Eisenberg , Yuliya Mishura

We consider dynamic pricing with covariates under a generalized linear demand model: a seller can dynamically adjust the price of a product over a horizon of $T$ time periods, and at each time period $t$, the demand of the product is…

Machine Learning · Computer Science 2023-11-14 Hanzhao Wang , Kalyan Talluri , Xiaocheng Li

This paper studies a general class of social choice problems in which agents' payoff functions (or types) are privately observable random variables, and monetary transfers are not available. We consider cardinal social choice functions…

Theoretical Economics · Economics 2024-08-20 Kazuya Kikuchi , Yukio Koriyama

Desharnais, Gupta, Jagadeesan and Panangaden introduced a family of behavioural pseudometrics for probabilistic transition systems. These pseudometrics are a quantitative analogue of probabilistic bisimilarity. Distance zero captures…

Logic in Computer Science · Computer Science 2015-07-01 Franck van Breugel , Babita Sharma , James Worrell

The undercut procedure was presented by Brams et al. [2] as a procedure for identifying an envy-free allocation when agents have preferences over sets of objects. They assumed that agents have strict preferences over objects and their…

Computer Science and Game Theory · Computer Science 2014-02-11 Haris Aziz

In Demand Response programs, price incentives might not be sufficient to modify residential consumers load profile. Here, we consider that each consumer has a preferred profile and a discomfort cost when deviating from it. Consumers can…

Optimization and Control · Mathematics 2017-12-01 Paulin Jacquot , Olivier Beaude , Nadia Oudjane , Stephane Gaubert