English
Related papers

Related papers: Personalized Pricing with Invalid Instrumental Var…

200 papers

Mean estimation under differential privacy is a fundamental problem, but worst-case optimal mechanisms do not offer meaningful utility guarantees in practice when the global sensitivity is very large. Instead, various heuristics have been…

Cryptography and Security · Computer Science 2021-11-02 Ziyue Huang , Yuting Liang , Ke Yi

Learning customer preferences from an observed behaviour is an important topic in the marketing literature. Structural models typically model forward-looking customers or firms as utility-maximizing agents whose utility is estimated using…

Computational Finance · Quantitative Finance 2017-12-14 Igor Halperin

This paper considers the optimal portfolio selection problem in a dynamic multi-period stochastic framework with regime switching. The risk preferences are of exponential (CARA) type with an absolute coefficient of risk aversion which…

Optimization and Control · Mathematics 2011-02-25 Traian A Pirvu , Huayue Zhang

Demand response (DR) has been demonstrated to be an effective method for reducing peak load and mitigating uncertainties on both the supply and demand sides of the electricity market. One critical question for DR research is how to…

Machine Learning · Computer Science 2023-06-27 Jun Song , Chaoyue Zhao

This paper develops an empirical balancing approach for the estimation of treatment effects under two-sided noncompliance using a binary conditionally independent instrumental variable. The method weighs both treatment and outcome…

Econometrics · Economics 2020-07-10 Phillip Heiler

We consider a revenue maximization model, in which a company aims at designing a menu of contracts, given a population of customers. A standard approach consists in constructing an incentive-compatible continuum of contracts, i.e., a menu…

Optimization and Control · Mathematics 2023-03-31 Quentin Jacquet , Wim van Ackooij , Clémence Alasseur , Stéphane Gaubert

The aim of inverse reinforcement learning (IRL) is to infer an agent's preferences from observing their behaviour. Usually, preferences are modelled as a reward function, $R$, and behaviour is modelled as a policy, $\pi$. One of the central…

Machine Learning · Computer Science 2024-12-17 Joar Skalse , Alessandro Abate

Offline Preference-based Reinforcement Learning (PbRL) learns rewards and policies aligned with human preferences without the need for extensive reward engineering and direct interaction with human annotators. However, ensuring safety…

Artificial Intelligence · Computer Science 2025-12-24 Ze Gong , Pradeep Varakantham , Akshat Kumar

Estimating the Individual Treatment Effect from observational data, defined as the difference between outcomes with and without treatment or intervention, while observing just one of both, is a challenging problems in causal learning. In…

Machine Learning · Computer Science 2020-05-07 Céline Beji , Michaël Bon , Florian Yger , Jamal Atif

We revisit the role of instrumental value as a driver of adaptive behavior. In active inference, instrumental or extrinsic value is quantified by the information-theoretic surprisal of a set of observations measuring the extent to which…

Neurons and Cognition · Quantitative Biology 2020-10-14 Alvaro Ovalle , Simon M. Lucas

Certified machine unlearning can be achieved via noise injection leading to differential privacy guarantees, where noise is calibrated to worst-case sensitivity. Such conservative calibration often results in performance degradation,…

Machine Learning · Computer Science 2026-05-25 Hanna Benarroch , Jamal Atif , Olivier Cappé

Selecting the best regularization parameter in inverse problems is a classical and yet challenging problem. Recently, data-driven approaches have become popular to tackle this challenge. These approaches are appealing since they do require…

Statistics Theory · Mathematics 2025-10-22 Jonathan Chirinos Rodriguez , Ernesto De Vito , Cesare Molinari , Lorenzo Rosasco , Silvia Villa

In survival contexts, substantial literature exists on estimating optimal treatment regimes, where treatments are assigned based on personal characteristics to maximize the survival probability. These methods assume that a set of covariates…

Methodology · Statistics 2025-07-24 Junwen Xia , Zishu Zhan , Jingxiao Zhang

The method of instrumental variables provides a fundamental and practical tool for causal inference in many empirical studies where unmeasured confounding between the treatments and the outcome is present. Modern data such as the genetical…

Methodology · Statistics 2022-10-28 Ziang Niu , Yuwen Gu , Wei Li

This paper studies the design of an optimal privacyaware estimator of a public random variable based on noisy measurements which contain private information. The public random variable carries non-private information, however, its estimate…

Optimization and Control · Mathematics 2018-08-08 Ehsan Nekouei , Henrik Sandberg , Mikael Skoglund , Karl H. Johansson

We propose and implement an approach to inference in linear instrumental variables models which is simultaneously robust and computationally tractable. Inference is based on self-normalization of sample moment conditions, and allows for…

Econometrics · Economics 2022-11-29 Eric Gautier , Christiern Rose

The instrumental variables (IV) method is a method for making causal inferences about the effect of a treatment based on an observational study in which there are unmeasured confounding variables. The method requires a valid IV, a variable…

Methodology · Statistics 2014-08-19 Dylan Small , Zhiqiang Tan , Scott Lorch , Alan Brookhart

In incomplete financial markets, pricing and hedging European options lack a unique no-arbitrage solution due to unhedgeable risks. This paper introduces a constrained deep learning approach to determine option prices and hedging strategies…

Computational Finance · Quantitative Finance 2025-11-27 Nicolas Baradel

We study an off-policy contextual pricing problem where the seller has access to samples of prices that customers were previously offered, whether they purchased at that price, and auxiliary features describing the customer and/or item…

Machine Learning · Computer Science 2023-01-12 Max Biggs

Offline reinforcement learning (RL) aims to find optimal policies in dynamic environments in order to maximize the expected total rewards by leveraging pre-collected data. Learning from heterogeneous data is one of the fundamental…

Machine Learning · Statistics 2026-03-10 Rui Miao , Babak Shahbaba , Annie Qu