Related papers: Entropy-Regularized Mean-Variance Portfolio Optimi…
Modelling extreme events and heavy-tailed phenomena is central to building reliable predictive systems in domains such as finance, climate science, and safety-critical AI. While L\'evy processes provide a natural mathematical framework for…
In this paper, we study a linear-quadratic optimal control problem for mean-field stochastic differential equations driven by a Poisson random martingale measure and a multidimensional Brownian motion. Firstly, the existence and uniqueness…
We consider the problem of utility maximization with exponential preferences in a market where the traded stock/risky asset price is modelled as a L\'evy-driven pure jump process (i.e. the driving L\'evy process has no Brownian component).…
Multi-period mean-variance optimization is a long-standing problem, caused by the failure of dynamic programming principle. This paper studies the mean-variance optimization in a setting of finite-horizon discrete-time Markov decision…
We introduce a new probabilistic method for solving a class of impulse control problems based on their representations as Backward Stochastic Differential Equations (BSDEs for short) with constrained jumps. As an example, our method is used…
To improve the efficient frontier of the classical mean-variance model in continuous time, we propose a varying terminal time mean-variance model with a constraint on the mean value of the portfolio asset, which moves with the varying…
We consider a couple of integrodifferential PDEs arising from a stochastic Markovian control problem subjected to initial-terminal conditions. These equations correspond to the MFG system for a controlled jump-diffusion process. We prove…
We analyze the consumption-portfolio selection problem of an investor facing both Brownian and jump risks. We bring new tools, in the form of orthogonal decompositions, to bear on the problem in order to determine the optimal portfolio in…
This paper explores decentralized learning in a graph-based setting, where data is distributed across nodes. We investigate a decentralized SGD algorithm that utilizes a random walk to update a global model based on local data. Our focus is…
This paper studies a robust stochastic control problem with a monotone mean-variance cost functional and random coefficients. The main technique is to find the saddle point through two backward stochastic differential equations (BSDEs) with…
Choosing a portfolio of risky assets over time that maximizes the expected return at the same time as it minimizes portfolio risk is a classical problem in Mathematical Finance and is referred to as the dynamic Markowitz problem (when the…
This paper investigates a continuous-time portfolio optimization problem with the following features: (i) a no-short selling constraint; (ii) a leverage constraint, that is, an upper limit for the sum of portfolio weights; and (iii) a…
This paper is concerned with a stochastic linear-quadratic optimal control problem with regime switching, random coefficients, and cone control constraint. The randomness of the coefficients comes from two aspects: the Brownian motion and…
We consider an expected utility maximization problem where the utility function is not necessarily concave and the time horizon is uncertain. We establish a necessary and sufficient condition for the optimality for general non-concave…
This paper studies the monotone mean-variance (MMV) problem and the classical mean-variance (MV) problem with convex cone trading constraints in a market with random coefficients. We provide semiclosed optimal strategies and optimal values…
In this paper, we study the optimal capital structure model with endogenous bankruptcy when the firm's asset value follows an exponential L\'evy process with positive jumps. In the Leland-Toft framework \cite{LelandToft96}, we obtain the…
Entropy regularization is known to improve exploration in sequential decision-making problems. We show that this same mechanism can also lead to nearly unbiased and lower-variance estimates of the mean reward in the optimize-and-estimate…
In this paper we derive a novel characterization result for time-consistent stochastic control problems with higher-order moments, originally formulated by Wang et al. [SIAM J. Control. Optim., 63 (2025), 1560--1589], and newly explore many…
We study the optimal dividend problem in the dual model where dividend payments can only be made at the jump times of an independent Poisson process. In this context, Avanzi et al. [5] solved the case with i.i.d. hyperexponential jumps;…
Modern recommendation systems rely on exploration to learn user preferences for new items, typically implementing uniform exploration policies (e.g., epsilon-greedy) due to their simplicity and compatibility with machine learning (ML)…