English
Related papers

Related papers: Variable Annuity with GMWB: surrender or not, that…

200 papers

In this paper, an optimization problem for the monotone mean-variance(MMV) criterion is considered in the perspective of the insurance company. The MMV criterion is an amended version of the classical mean-variance(MV) criterion which…

Optimization and Control · Mathematics 2022-12-05 Bohan Li , Junyi Guo , Linlin Tian

In a multi-armed bandit (MAB) problem a gambler needs to choose at each round of play one of K arms, each characterized by an unknown reward distribution. Reward realizations are only observed when an arm is selected, and the gambler's…

Machine Learning · Computer Science 2019-06-11 Omar Besbes , Yonatan Gur , Assaf Zeevi

Money-back guarantees (MBGs) are features of pooled retirement income products that address bequest concerns by ensuring the initial premium is returned through lifetime payments or, upon early death, as a death benefit to the estate. This…

Portfolio Management · Quantitative Finance 2026-02-19 German Nova Orozco , Duy-Minh Dang , Peter A. Forsyth

We introduce a novel extension of the canonical multi-armed bandit problem that incorporates an additional strategic innovation: abstention. In this enhanced framework, the agent is not only tasked with selecting an arm at each time step,…

Machine Learning · Computer Science 2026-03-24 Junwen Yang , Tianyuan Jin , Vincent Y. F. Tan

Algorithms for the Multi-Armed Bandit (MAB) problem play a central role in sequential decision-making and have been extensively explored both theoretically and numerically. While most classical approaches aim to identify the arm with the…

Machine Learning · Computer Science 2026-04-02 Gabriel Turinici

Motivated by recommendation problems in music streaming platforms, we propose a nonstationary stochastic bandit model in which the expected reward of an arm depends on the number of rounds that have passed since the arm was last pulled.…

Machine Learning · Statistics 2020-02-20 Leonardo Cella , Nicolò Cesa-Bianchi

Minimizing volatility and adjustment costs is of central importance in many economic environments, yet it is often complicated by evolving feasibility constraints. We study a decision maker who repeatedly selects an action from a…

Theoretical Economics · Economics 2026-02-18 Simon Jantschgi , Heinrich H. Nax , Bary S. R. Pradelski , Marek Pycia

When selling many goods with independent valuations, we develop a distributionally robust framework, consisting of a two-player game between seller and nature. The seller has only limited knowledge about the value distribution. The seller…

Computer Science and Game Theory · Computer Science 2026-03-30 Tim S. G. van Eck , Pieter Kleer , Johan S. H. van Leeuwaarden

We study the optimal investment-consumption problem for a member of defined contribution plan during the decumulation phase. For a fixed annuitization time, to achieve higher final annuity, we consider a variable consumption rate. Moreover,…

Portfolio Management · Quantitative Finance 2020-08-18 Hassan Dadashi

We study the problem of \emph{dynamic regret minimization} in $K$-armed Dueling Bandits under non-stationary or time varying preferences. This is an online learning setup where the agent chooses a pair of items at each round and observes…

Machine Learning · Computer Science 2022-06-14 Aadirupa Saha , Shubham Gupta

We study an optimal dividend problem under a bankruptcy constraint. Firms face a trade-off between potential bankruptcy and extraction of profits. In contrast to previous works, general cash flow drifts, including Ornstein--Uhlenbeck and…

Optimization and Control · Mathematics 2018-03-05 Max Reppen , Jean-Charles Rochet , H. Mete Soner

A survey is performed of various Multi-Armed Bandit (MAB) strategies in order to examine their performance in circumstances exhibiting non-stationary stochastic reward functions in conjunction with delayed feedback. We run several MAB…

Machine Learning · Computer Science 2019-07-31 Larkin Liu , Richard Downe , Joshua Reid

In today's online advertising markets, a crucial requirement for an advertiser is to control her total expenditure within a time horizon under some budget. Among various budget control methods, throttling has emerged as a popular choice,…

Computer Science and Game Theory · Computer Science 2023-12-14 Zhaohua Chen , Chang Wang , Qian Wang , Yuqi Pan , Zhuming Shi , Zheng Cai , Yukun Ren , Zhihua Zhu , Xiaotie Deng

In a combinatorial auction with item bidding, agents participate in multiple single-item second-price auctions at once. As some items might be substitutes, agents need to strategize in order to maximize their utilities. A number of results…

Computer Science and Game Theory · Computer Science 2017-04-17 Paul Dütting , Thomas Kesselheim

When multi-armed bandit (MAB) algorithms allocate pulls among competing arms, the resulting allocation can exhibit huge variation. This is particularly harmful in modern applications such as learning-enhanced platform operations and…

Machine Learning · Computer Science 2026-02-10 Yilun Chen , Jiaqi Lu

In feature-based dynamic pricing, a seller sets appropriate prices for a sequence of products (described by feature vectors) on the fly by learning from the binary outcomes of previous sales sessions ("Sold" if valuation $\geq$ price, and…

Machine Learning · Computer Science 2022-04-04 Jianyu Xu , Yu-Xiang Wang

In this paper we study a generalized version of classical multi-armed bandits (MABs) problem by allowing for arbitrary constraints on constituent bandits at each decision point. The motivation of this study comes from many situations that…

Machine Learning · Computer Science 2014-10-07 Xiang-yang Li , Shaojie Tang , Yaqin Zhou

In this paper, a robust optimal reinsurance-investment problem with delay is studied under the $\alpha$-maxmin mean-variance criterion. The surplus process of an insurance company approximates Brownian motion with drift. The financial…

Optimization and Control · Mathematics 2022-09-13 Min Zhang , Yong He

We consider a Markovian single server queue with impatient customers. There is a customer abandonment cost and a holding cost for customers in the system. We consider two versions of the problem. In the first version, customers pay a reward…

Optimization and Control · Mathematics 2025-11-12 Runhua Wu , Hayriye Ayhan

The present paper addresses the issue of the stochastic control of the optimal dynamic reinsurance policy and dynamic dividend strategy, which are state-dependent, for an insurance company that operates under multiple insurance lines of…

Optimization and Control · Mathematics 2020-02-11 Khaled Masoumifard , Mohammad Zokaei