English
Related papers

Related papers: Variable Annuity with GMWB: surrender or not, that…

200 papers

Offline reinforcement learning and offline inverse reinforcement learning aim to recover near-optimal value functions or reward models from a fixed batch of logged trajectories, yet current practice still struggles to enforce Bellman…

Machine Learning · Computer Science 2026-01-27 Enoch H. Kang , Kyoungseok Jang

In this paper, we study a mean-variance optimization problem in an infinite horizon discrete time discounted Markov decision process (MDP). The objective is to minimize the variance of system rewards with the constraint of mean performance.…

Optimization and Control · Mathematics 2017-08-24 Li Xia

This paper considers the multi-armed bandit (MAB) problem and provides a new best-of-both-worlds (BOBW) algorithm that works nearly optimally in both stochastic and adversarial settings. In stochastic settings, some existing BOBW algorithms…

Machine Learning · Computer Science 2022-06-15 Shinji Ito , Taira Tsuchiya , Junya Honda

In this article we study an optimal stopping/optimal control problem which models the decision facing a risk-averse agent over when to sell an asset. The market is incomplete so that the asset exposure cannot be hedged. In addition to the…

Portfolio Management · Quantitative Finance 2008-12-10 Vicky Henderson , David Hobson

This study considers an optimal reinsurance, investment, and dividend strategy control problem for insurance companies in a regulated Markov regime-switching environment, intending to maximize long-run average reward. Unlike existing single…

Optimization and Control · Mathematics 2025-12-18 Lingjia Zeng , Manman Li

We study the problem of fair online resource allocation via non-monetary mechanisms, where multiple agents repeatedly share a resource without monetary transfers. Previous work has shown that every agent can guarantee $1/2$ of their ideal…

Computer Science and Game Theory · Computer Science 2025-05-27 David X. Lin , Daniel Hall , Giannis Fikioris , Siddhartha Banerjee , Éva Tardos

This paper studies the valuation and optimal strategy of convertible bonds as a Dynkin game by using the reflected backward stochastic differential equation method and the variational inequality method. We first reduce such a Dynkin game to…

Mathematical Finance · Quantitative Finance 2015-04-01 Huiwen Yan , Zhou Yang , Fahuai Yi , Gechun Liang

The system operator's scheduling problem in electricity markets, called unit commitment, is a non-convex mixed-integer program. The optimal value function is non-convex, preventing the application of traditional marginal pricing theory to…

General Economics · Economics 2024-10-03 Conleigh Byers , Brent Eldridge

Experimentation with interference poses a significant challenge in contemporary online platforms. Prior research on experimentation with interference has concentrated on the final output of a policy. The cumulative performance, while…

Machine Learning · Computer Science 2024-07-17 Su Jia , Peter Frazier , Nathan Kallus

We study the piecewise stationary combinatorial semi-bandit problem with causally related rewards. In our nonstationary environment, variations in the base arms' distributions, causal relationships between rewards, or both, change the…

Machine Learning · Computer Science 2023-07-27 Behzad Nourani-Koliji , Steven Bilaj , Amir Rezaei Balef , Setareh Maghsudi

We study the problem of pricing variable annuities with a multi-layer expense strategy, under which the insurer charges fees from the policyholder's account only when the account value lies in some pre-specified disjoint intervals, where on…

Probability · Mathematics 2015-12-14 Jiang Zhou , Lan Wu

This paper studies an optimal dividend problem for a company that aims to maximize the mean-variance (MV) objective of the accumulated discounted dividend payments up to its ruin time. The MV objective involves an integral form over a…

Optimization and Control · Mathematics 2025-08-19 Jingyi Cao , Dongchen Li , Virginia R. Young , Bin Zou

We propose two parametric approaches to evaluate swing contracts with firm constraints. Our objective is to define approximations for the optimal control, which represents the amounts of energy purchased throughout the contract. The first…

Mathematical Finance · Quantitative Finance 2024-06-13 Vincent Lemaire , Gilles Pagès , Christian Yeo

We study online bilateral trade, where a learner facilitates repeated exchanges between a buyer and a seller to maximize the Gain From Trade (GFT), i.e., the social welfare. In doing so, the learner must guarantee not to subsidize the…

Computer Science and Game Theory · Computer Science 2026-02-06 Anna Lunghi , Mattia Piccinato , Matteo Castiglioni , Alberto Marchesi

We study non-rectangular robust Markov decision processes under the average-reward criterion, where the ambiguity set couples transition probabilities across states and the adversary commits to a stationary kernel for the entire horizon. We…

Optimization and Control · Mathematics 2026-03-11 Shengbo Wang , Nian Si

This paper explores the application of Machine Learning techniques for pricing high-dimensional options within the framework of the Uncertain Volatility Model (UVM). The UVM is a robust framework that accounts for the inherent…

Computational Finance · Quantitative Finance 2025-06-06 Ludovic Goudenege , Andrea Molent , Antonino Zanette

In this paper we consider some insurance policies related to drawdown and drawup events of log-returns for an underlying asset modeled by a spectrally negative geometric L\'evy process. We consider four contracts, three of which were…

Pricing of Securities · Quantitative Finance 2017-10-10 Zbigniew Palmowski , Joanna Tumilewicz

We consider the classical mathematical economics problem of {\em Bayesian optimal mechanism design} where a principal aims to optimize expected revenue when allocating resources to self-interested agents with preferences drawn from a known…

Computer Science and Game Theory · Computer Science 2010-01-15 Shuchi Chawla , Jason Hartline , David Malec , Balasubramanian Sivan

We study the stochastic Budgeted Multi-Armed Bandit (MAB) problem, where a player chooses from $K$ arms with unknown expected rewards and costs. The goal is to maximize the total reward under a budget constraint. A player thus seeks to…

Machine Learning · Computer Science 2023-08-16 Marco Heyden , Vadim Arzamasov , Edouard Fouché , Klemens Böhm

We study a variant of the classical stochastic $K$-armed bandit where observing the outcome of each arm is expensive, but cheap approximations to this outcome are available. For example, in online advertising the performance of an ad can be…

Machine Learning · Computer Science 2016-11-01 Kirthevasan Kandasamy , Gautam Dasarathy , Jeff Schneider , Barnabás Póczos