English
Related papers

Related papers: Inference of Utilities and Time Preference in Sequ…

200 papers

We study active preference learning as a framework for intuitively specifying the behaviour of autonomous robots. In active preference learning, a user chooses the preferred behaviour from a set of alternatives, from which the robot learns…

Robotics · Computer Science 2020-09-30 Nils Wilde , Dana Kulic , Stephen L. Smith

This paper studies temporal planning in probabilistic environments, modeled as labeled Markov decision processes (MDPs), with user preferences over multiple temporal goals. Existing works reflect such preferences as a prioritized list of…

Formal Languages and Automata Theory · Computer Science 2023-04-25 Lening Li , Hazhar Rahmani , Jie Fu

In this paper we study the optimal investment and reinsurance problem of an insurance company whose investment preferences are described via a forward dynamic exponential utility in a regime-switching market model. Financial and actuarial…

Portfolio Management · Quantitative Finance 2021-06-29 Katia Colaneri , Alessandra Cretarola , Benedetta Salterini

Human preferences are not always represented via complete linear orders: It is natural to employ partially-ordered preferences for expressing incomparable outcomes. In this work, we consider decision-making and probabilistic planning in…

Robotics · Computer Science 2024-10-21 Hazhar Rahmani , Abhishek N. Kulkarni , Jie Fu

Caching algorithms try to predict content popularity, and place the content closer to the users. Additionally, nowadays requests are increasingly driven by recommendation systems (RS). These important trends, point to the following:…

Networking and Internet Architecture · Computer Science 2021-10-05 Theodoros Giannakas , Pavlos Sermpezis , Thrasyvoulos Spyropoulos

We consider an illiquid financial market with different regimes modeled by a continuous-time finite-state Markov chain. The investor can trade a stock only at the discrete arrival times of a Cox process with intensity depending on the…

Portfolio Management · Quantitative Finance 2012-04-26 Paul Gassiat , Fausto Gozzi , Huyên Pham

User interests are usually dynamic in the real world, which poses both theoretical and practical challenges for learning accurate preferences from rich behavior data. Among existing user behavior modeling solutions, attention networks are…

Information Retrieval · Computer Science 2022-04-14 Chao Chen , Haoyu Geng , Nianzu Yang , Junchi Yan , Daiyue Xue , Jianping Yu , Xiaokang Yang

In this paper we present a framework for risk-sensitive model predictive control (MPC) of linear systems affected by stochastic multiplicative uncertainty. Our key innovation is to consider a time-consistent, dynamic risk evaluation of the…

Optimization and Control · Mathematics 2018-04-26 Sumeet Singh , Yin-Lam Chow , Anirudha Majumdar , Marco Pavone

The Random Utility Maximization model is by far the most adopted framework to estimate consumer choice behavior. However, behavioral economics has provided strong empirical evidence of irrational choice behavior, such as halo effects, that…

Econometrics · Economics 2021-09-10 Sanjay Dominik Jena , Andrea Lodi , Claudio Sole

We analyze the problem of learning a single user's preferences in an active learning setting, sequentially and adaptively querying the user over a finite time horizon. Learning is conducted via choice-based queries, where the user selects…

Machine Learning · Statistics 2017-02-27 Stephen N. Pallone , Peter I. Frazier , Shane G. Henderson

The aim of this work consists in the study of the optimal investment strategy for a behavioural investor, whose preference towards risk is described by both a probability distortion and an S-shaped utility function. Within a continuous-time…

Portfolio Management · Quantitative Finance 2013-04-30 Miklos Rasonyi , Andrea M. Rodrigues

Managing stock efficiently remains a core issue in modern logistics, where companies must reconcile cost efficiency with dependable service despite unpredictable market conditions. Conventional models often overlook the direct connection…

Optimization and Control · Mathematics 2026-04-14 Tianxiao Sun , Noah Schwarzkopf

Markov automata combine non-determinism, probabilistic branching, and exponentially distributed delays. This compositional variant of continuous-time Markov decision processes is used in reliability engineering, performance evaluation and…

Logic in Computer Science · Computer Science 2017-05-11 Tim Quatmann , Sebastian Junges , Joost-Pieter Katoen

This study develops an inverse portfolio optimization framework for recovering latent investor preferences including risk aversion, transaction cost sensitivity, and ESG orientation from observed portfolio allocations. Using controlled…

General Finance · Quantitative Finance 2025-10-14 Jinho Cha , Long Pham , Thi Le Hoa Vo , Jaeyoung Cho , Jaejin Lee

Given a sequence of sets, where each set has a timestamp and contains an arbitrary number of elements, temporal sets prediction aims to predict the elements in the subsequent set. Previous studies for temporal sets prediction mainly focus…

Machine Learning · Computer Science 2023-08-29 Le Yu , Zihang Liu , Leilei Sun , Bowen Du , Chuanren Liu , Weifeng Lv

We investigate a continuous-time investment-consumption problem with model uncertainty in a general diffusion-based market with random model coefficients. We assume that a power utility investor is ambiguity-averse, with the preference to…

Portfolio Management · Quantitative Finance 2024-07-04 Len Patrick Dominic M. Garces , Yang Shen

We derive a family of risk-sensitive reinforcement learning methods for agents, who face sequential decision-making tasks in uncertain environments. By applying a utility function to the temporal difference (TD) error, nonlinear…

Machine Learning · Computer Science 2014-10-10 Yun Shen , Michael J. Tobia , Tobias Sommer , Klaus Obermayer

Recommender systems are widely used for suggesting books, education materials, and products to users by exploring their behaviors. In reality, users' preferences often change over time, leading to studies on time-dependent recommender…

Information Retrieval · Computer Science 2024-12-17 Haidong Zhang , Wancheng Ni , Xin Li , Yiping Yang

In intertemporal settings, the multiattribute utility theory of Kihlstrom and Mirman suggests the application of a concave transform of the lifetime utility index. This construction, while allowing time and risk attitudes to be separated,…

Mathematical Finance · Quantitative Finance 2024-10-07 Luca De Gennaro Aquino , Sascha Desmettre , Yevhen Havrylenko , Mogens Steffensen

In this paper, we propose a machine learning algorithm for time-inconsistent portfolio optimization. The proposed algorithm builds upon neural network based trading schemes, in which the asset allocation at each time point is determined by…

Portfolio Management · Quantitative Finance 2023-09-06 Kristoffer Andersson , Cornelis W. Oosterlee