English
Related papers

Related papers: Discrete time multi-period mean-variance model: Be…

200 papers

We consider the problem of finding optimal time-periodic sensor schedules for estimating the state of discrete-time dynamical systems. We assume that {multiple} sensors have been deployed and that the sensors are subject to resource…

Applications · Statistics 2016-11-17 Sijia Liu , Makan Fardad , Engin Masazade , Pramod K. Varshney

We study time-inconsistent recursive stochastic control problems, i.e., for which the Bellman principle of optimality does not hold. For this class of problems classical optimal controls may fail to exist, or to be relevant in practice, and…

Optimization and Control · Mathematics 2024-03-14 Elisa Mastrogiacomo , Marco Tarsia

We propose and study a simple model of dynamical redistribution of capital in a diversified portfolio. We consider a hypothetical situation of a portfolio composed of N uncorrelated stocks. Each stock price follows a multiplicative random…

Statistical Mechanics · Physics 2015-06-25 Matteo Marsili , Sergei Maslov , Yi-Cheng Zhang

In this paper, we investigate the effects of applying generalised (non-exponential) discounting on a long-run impulse control problem for a Feller-Markov process. We show that the optimal value of the discounted problem is the same as the…

Optimization and Control · Mathematics 2024-04-22 Damian Jelito , Łukasz Stettner

This paper is concerned with an optimal reinsurance and investment problem for an insurance firm under the criterion of mean-variance. The driving Brownian motion and the rate in return of the risky asset price dynamic equation cannot be…

Optimization and Control · Mathematics 2020-06-04 Shihao Zhu , Jingtao Shi

Optimization of decision problems in stochastic environments is usually concerned with maximizing the probability of achieving the goal and minimizing the expected episode length. For interacting agents in time-critical applications,…

Artificial Intelligence · Computer Science 2007-05-23 Balint Takacs , Istvan Szita , Andras Lorincz

We consider the stochastic optimal control problem of McKean-Vlasov stochastic differential equation where the coefficients may depend upon the joint law of the state and control. By using feedback controls, we reformulate the problem into…

Probability · Mathematics 2017-03-09 Huyên Pham , Xiaoli Wei

This is a companion paper of [Mixed equilibrium solution of time-inconsistent stochastic LQ problem, arXiv:1802.03032], where general theory has been established to characterize the open-loop equilibrium control, feedback equilibrium…

Optimization and Control · Mathematics 2018-03-26 Yuan-Hua Ni , Xun Li , Ji-Feng Zhang , Miroslav Krstic

We study equilibrium feedback strategies for a family of dynamic mean-variance problems with competition among a large group of agents. We assume that the time horizon is random and each agent's risk aversion depends dynamically on the…

Optimization and Control · Mathematics 2026-05-05 Xiaoqing Liang , Jie Xiong , Ying Yang

In this paper, we propose a machine learning algorithm for time-inconsistent portfolio optimization. The proposed algorithm builds upon neural network based trading schemes, in which the asset allocation at each time point is determined by…

Portfolio Management · Quantitative Finance 2023-09-06 Kristoffer Andersson , Cornelis W. Oosterlee

While maximizing expected return is the goal in most reinforcement learning approaches, risk-sensitive objectives such as conditional value at risk (CVaR) are more suitable for many high-stakes applications. However, relatively little is…

Machine Learning · Computer Science 2020-04-06 Ramtin Keramati , Christoph Dann , Alex Tamkin , Emma Brunskill

We study finite-horizon continuous-time policy evaluation from discrete closed-loop trajectories under time-inhomogeneous dynamics. The target value surface solves a backward parabolic equation, but the Bellman baseline obtained from…

Machine Learning · Statistics 2026-05-11 Yaowei Zheng , Richong Zhang , Shenxi Wu , Shirui Bian , Haosong Zhang , Li Zeng , Xingjian Ma , Yichi Zhang

Regime-switching poses both problems and opportunities for portfolio managers. If a switch in the behaviour of the markets is not quickly detected it can be a source of loss, since previous trading positions may be inappropriate in the new…

Computational Engineering, Finance, and Science · Computer Science 2023-08-21 Piotr Pomorski , Denise Gorse

The paper studies problem of continuous time optimal portfolio selection for a incom- plete market diffusion model. It is shown that, under some mild conditions, near optimal strategies for investors with different performance criteria can…

Portfolio Management · Quantitative Finance 2014-04-15 Nikolai Dokuchaev

Estimating dynamic treatment regimes (DTRs) from retrospective observational data is challenging as some degree of unmeasured confounding is often expected. In this work, we develop a framework of estimating properly defined "optimal" DTRs…

Methodology · Statistics 2021-04-19 Shuxiao Chen , Bo Zhang

We investigate a time-inconsistent, non-Markovian finite-player game in continuous time, where each player's objective functional depends non-linearly on the expected value of the state process. As a result, the classical Bellman optimality…

Probability · Mathematics 2025-12-10 Dylan Possamaï , Chiara Rossato

The dynamic allocation problem, also known as the `multi-armed bandit' problem, simulates a situation in which an agent is faced with a tradeoff between actions that yield an immediate reward and actions whose benefits can only be perceived…

Probability · Mathematics 2026-02-03 Christopher Wang

Differential Dynamic Programming is an optimal control technique often used for trajectory generation. Many variations of this algorithm have been developed in the literature, including algorithms for stochastic dynamics or state and input…

Optimization and Control · Mathematics 2022-05-26 Dennis Gramlich , Carsten W. Scherer , Christian Ebenbauer

This article introduces a novel dynamic framework to Bayesian model averaging for time-varying parameter quantile regressions. By employing sequential Markov chain Monte Carlo, we combine empirical estimates derived from dynamically chosen…

Statistics Theory · Mathematics 2024-11-08 Mauro Bernardi , Roberto Casarin , Bertrand Maillet , Lea Petrella

Current approaches to model-based offline reinforcement learning often incorporate uncertainty-based reward penalization to address the distributional shift problem. These approaches, commonly known as pessimistic value iteration, use Monte…

Machine Learning · Computer Science 2025-01-17 Abdullah Akgül , Manuel Haußmann , Melih Kandemir