English
Related papers

Related papers: Insurance Pricing Optimization via Off-Policy Eval…

200 papers

Off-policy estimation (OPE) methods enable unbiased offline evaluation of recommender systems, directly estimating the online reward some target policy would have obtained, from offline data and with statistical guarantees. The theoretical…

Machine Learning · Statistics 2025-08-12 Olivier Jeunen

The online advertising market, with its thousands of auctions run per second, presents a daunting challenge for advertisers who wish to optimize their spend under a budget constraint. Thus, advertising platforms typically provide automated…

Machine Learning · Computer Science 2023-10-17 Dmytro Korenkevych , Frank Cheng , Artsiom Balakir , Alex Nikulkov , Lingnan Gao , Zhihao Cen , Zuobing Xu , Zheqing Zhu

With the rapid rise of InsurTech, traditional insurance companies are increasingly exploring alternative data sources and advanced technologies to sustain their competitive edge. This paper provides both a conceptual overview and practical…

Computation and Language · Computer Science 2025-12-19 Panyi Dong , Zhiyu Quan

Multi-objective optimization is a type of decision making problems where multiple conflicting objectives are optimized. We study offline optimization of multi-objective policies from data collected by an existing policy. We propose a…

Machine Learning · Computer Science 2023-10-31 Shima Alizadeh , Aniruddha Bhargava , Karthick Gopalswamy , Lalit Jain , Branislav Kveton , Ge Liu

This paper considers an aggregator of Electric Vehicles (EVs) who aims to learn the aggregate power of his/her fleet while also participating in the electricity market. The proposed approach is based on a data-driven inverse optimization…

Systems and Control · Electrical Eng. & Systems 2021-03-08 Ricardo Fernández-Blanco , Juan Miguel Morales , Salvador Pineda , Álvaro Porras

We propose policy gradient algorithms for solving a risk-sensitive reinforcement learning (RL) problem in on-policy as well as off-policy settings. We consider episodic Markov decision processes, and model the risk using the broad class of…

Machine Learning · Computer Science 2024-06-25 Nithia Vijayan , Prashanth L. A

Propensity score methods are widely used for estimating treatment effects from observational studies. A popular approach is to estimate propensity scores by maximum likelihood based on logistic regression, and then apply inverse probability…

Methodology · Statistics 2017-10-24 Zhiqiang Tan

Price determination is a central research topic of revenue management in marketing. The important aspect in pricing is controlling the stochastic behavior of demand, and the previous studies have tackled price optimization problems with…

Optimization and Control · Mathematics 2024-01-04 Yuya Hikima , Akiko Takeda

In this paper we study the optimal investment and reinsurance problem of an insurance company whose investment preferences are described via a forward dynamic exponential utility in a regime-switching market model. Financial and actuarial…

Portfolio Management · Quantitative Finance 2021-06-29 Katia Colaneri , Alessandra Cretarola , Benedetta Salterini

This study models the monopoly pricing of weather index insurance as a Bowley-type sequential game involving a profit-maximizing insurer (leader) and a farmer (follower). The farmer chooses an insurance payoff to minimize a convex…

Risk Management · Quantitative Finance 2025-12-02 Tim J. Boonen , Wenyuan Li , Zixiao Quan

Primal-dual safe RL methods commonly perform iterations between the primal update of the policy and the dual update of the Lagrange Multiplier. Such a training paradigm is highly susceptible to the error in cumulative cost estimation since…

Machine Learning · Computer Science 2024-04-16 Zifan Wu , Bo Tang , Qian Lin , Chao Yu , Shangqin Mao , Qianlong Xie , Xingxing Wang , Dong Wang

We propose an estimator and confidence interval for computing the value of a policy from off-policy data in the contextual bandit setting. To this end we apply empirical likelihood techniques to formulate our estimator and confidence…

Machine Learning · Computer Science 2020-10-20 Nikos Karampatziakis , John Langford , Paul Mineiro

This paper considers the pricing of equity-linked life insurance contracts with death and survival benefits in a general model with multiple stochastic risk factors: interest rate, equity, volatility, unsystematic and systematic mortality.…

Pricing of Securities · Quantitative Finance 2021-11-03 Karim Barigou , Lukasz Delong

Off-policy policy optimization is a challenging problem in reinforcement learning (RL). The algorithms designed for this problem often suffer from high variance in their estimators, which results in poor sample efficiency, and have issues…

Machine Learning · Computer Science 2020-09-15 Daoming Lyu , Qi Qi , Mohammad Ghavamzadeh , Hengshuai Yao , Tianbao Yang , Bo Liu

At the core of insurance business lies classification between risky and non-risky insureds, actuarial fairness meaning that risky insureds should contribute more and pay a higher premium than non-risky or less-risky ones. Actuaries,…

Machine Learning · Statistics 2022-12-27 Vincent Grari , Arthur Charpentier , Marcin Detyniecki

We study the problem of off-policy evaluation (OPE) in Reinforcement Learning (RL), where the aim is to estimate the performance of a new policy given historical data that may have been generated by a different policy, or policies. In…

Machine Learning · Computer Science 2019-12-16 Aurélien F. Bibaut , Ivana Malenica , Nikos Vlassis , Mark J. van der Laan

This paper develops a risk-adjusted alternative to standard optimal policy learning (OPL) for observational data by importing Roy's (1952) safety-first principle into the treatment assignment problem. We formalize a welfare functional that…

Econometrics · Economics 2025-10-07 Giovanni Cerulli , Francesco Caracciolo

Optimal reinsurance when Value at Risk and expected surplus is balanced through their ratio is studied, and it is demonstrated how results for risk-adjusted surplus can be utilized. Simplifications for large portfolios are derived, and this…

Applications · Statistics 2019-12-10 Erik Bølviken , Yinzhi Wang

In this paper, we consider the problem of optimal reinsurance design, when the risk is measured by a distortion risk measure and the premium is given by a distortion risk premium. First, we show how the optimal reinsurance design for the…

Risk Management · Quantitative Finance 2014-06-12 Hirbod Assa

This contribution is concerned with price optimisation of the new business for a non-life product. Due to high competition in the insurance market, non-life insurers are interested in increasing their conversion rates on new business based…

Computational Finance · Quantitative Finance 2017-11-22 Maissa Tamraz , Yaming Yang