English
Related papers

Related papers: Variable Annuity with GMWB: surrender or not, that…

200 papers

The multi-armed bandit (MAB) model is one of the most classical models to study decision-making in an uncertain environment. In this model, a player chooses one of $K$ possible arms of a bandit machine to play at each time step, where the…

Machine Learning · Computer Science 2023-06-13 Bo Li , Chi Ho Yeung

We find the variance-optimal equivalent martingale measure when multivariate assets are modeled by a regime-switching geometric Brownian motion, and the regimes are represented by a homogeneous continuous time Markov chain. Under this new…

Probability · Mathematics 2023-09-14 Bruno Remillard , Sylvain Rubenthaler

Robust mechanism design is a rising alternative to Bayesian mechanism design, which yields designs that do not rely on assumptions like full distributional knowledge. We apply this approach to mechanisms for selling a single item, assuming…

Computer Science and Game Theory · Computer Science 2022-05-24 Nir Bachrach , Inbal Talgam-Cohen

We consider the problem of stochastic optimal control, where the state-feedback control policies take the form of a probability distribution and where a penalty on the entropy is added. By viewing the cost function as a Kullback- Leibler…

Optimization and Control · Mathematics 2024-12-12 Marc Lambert , Francis Bach , Silvère Bonnabel

Inspired by real-time ad exchanges for online display advertising, we consider the problem of inferring a buyer's value distribution for a good when the buyer is repeatedly interacting with a seller through a posted-price mechanism. We…

Machine Learning · Computer Science 2013-11-28 Kareem Amin , Afshin Rostamizadeh , Umar Syed

A combinatorial market consists of a set of indivisible items and a set of agents, where each agent has a valuation function that specifies for each subset of items its value for the given agent. From an optimization point of view, the goal…

Computer Science and Game Theory · Computer Science 2023-01-05 Kristóf Bérczi , Laura Codazzi , Julian Golak , Alexander Grigoriev

Withdrawal guarantees ensure the periodical deduction of a constant dollar-amount from a fund investment for a fixed number of periods. If the fund depletes before the last withdrawal, the guarantor has to finance the outstanding…

Pricing of Securities · Quantitative Finance 2015-03-20 Andreas Kunz

We consider a stochastic multi-armed bandit setting where reward must be actively queried for it to be observed. We provide tight lower and upper problem-dependent guarantees on both the regret and the number of queries. Interestingly, we…

Machine Learning · Computer Science 2022-10-28 Nadav Merlis , Yonathan Efroni , Shie Mannor

This paper explores optimal insurance solutions based on the Lambda-Value-at-Risk ($\Lambda\VaR$). If the expected value premium principle is used, our findings confirm that, similar to the VaR model, a truncated stop-loss indemnity is…

Risk Management · Quantitative Finance 2025-08-19 Tim J. Boonen , Yuyu Chen , Xia Han , Qiuqi Wang

We study the stochastic multi-armed bandit (MAB) problem in the presence of side-observations across actions that occur as a result of an underlying network structure. In our model, a bipartite graph captures the relationship between…

Machine Learning · Computer Science 2017-07-14 Swapna Buccapatnam , Fang Liu , Atilla Eryilmaz , Ness B. Shroff

We investigate an optimal reinsurance problem for an insurance company facing a constant fixed cost when the reinsurance contract is signed. The insurer needs to optimally choose both the starting time of the reinsurance contract and the…

Mathematical Finance · Quantitative Finance 2021-01-14 Matteo Brachetta , Claudia Ceci

Stochastic gradient descent is the method of choice for large-scale machine learning problems, by virtue of its light complexity per iteration. However, it lags behind its non-stochastic counterparts with respect to the convergence rate,…

Machine Learning · Statistics 2016-03-23 Vatsal Shah , Megasthenis Asteris , Anastasios Kyrillidis , Sujay Sanghavi

We consider dynamic pricing with covariates under a generalized linear demand model: a seller can dynamically adjust the price of a product over a horizon of $T$ time periods, and at each time period $t$, the demand of the product is…

Machine Learning · Computer Science 2023-11-14 Hanzhao Wang , Kalyan Talluri , Xiaocheng Li

We introduce algorithms that achieve state-of-the-art \emph{dynamic regret} bounds for non-stationary linear stochastic bandit setting. It captures natural applications such as dynamic pricing and ads allocation in a changing environment.…

Machine Learning · Computer Science 2021-07-20 Wang Chi Cheung , David Simchi-Levi , Ruihao Zhu

In this paper we address the problem of optimal dividend payout strategies from a surplus process governed by Brownian motion with drift under a drawdown constraint, i.e. the dividend rate can never decrease below a given fraction $a$ of…

Optimization and Control · Mathematics 2022-06-27 Hansjoerg Albrecher , Pablo Azcue , Nora Muler

We propose an analytically tractable variation of the minority game in which rational agents use probabilistic strategies. In our model, $N$ agents choose between two alternatives repeatedly, and those who are in the minority get a pay-off…

Trading and Market Microstructure · Quantitative Finance 2014-02-11 V. Sasidevan , Deepak Dhar

In this paper we study the problem of optimal dividend payment strategy which maximizes the expected discounted sum of dividends to a multidimensional set up of n associated insurance companies where the surplus process follows an…

Optimization and Control · Mathematics 2018-10-04 Pablo Azcue , Nora Muler

This paper studies a class of constrained restless multi-armed bandits (CRMAB). The constraints are in the form of time varying set of actions (set of available arms). This variation can be either stochastic or semi-deterministic. Given a…

Systems and Control · Computer Science 2021-09-07 Kesav Kaza , Rahul Meshram , Varun Mehta , S. N. Merchant

Multi-armed bandit (MAB) is a widely adopted framework for sequential decision-making under uncertainty. Traditional bandit algorithms rely solely on online data, which tends to be scarce as it must be gathered during the online phase when…

Statistics Theory · Mathematics 2026-04-23 Wenlong Ji , Yihan Pan , Ruihao Zhu , Lihua Lei

Multi-armed bandit models have proven to be useful in modeling many real world problems in the areas of control and sequential decision making with partial information. However, in many scenarios, such as those prevalent in healthcare and…

Optimization and Control · Mathematics 2024-08-27 Qinyang He , Yonatan Mintz