English
Related papers

Related papers: An Evaluation Framework for Personalization Strate…

200 papers

We consider an online estimation problem involving a set of agents. Each agent has access to a (personal) process that generates samples from a real-valued distribution and seeks to estimate its mean. We study the case where some of the…

Machine Learning · Computer Science 2022-12-20 Mahsa Asadi , Aurélien Bellet , Odalric-Ambrym Maillard , Marc Tommasi

Bayesian experimental design (BED) is a principled framework for data-efficient design of sequential experiments. However, existing BED methods are unable to adapt to dynamic constraints inherent in real-world tasks due to budget…

Machine Learning · Statistics 2026-05-27 Yujia Guo , Daolang Huang , Xinyu Zhang , Sammie Katt , Samuel Kaski , Ayush Bharti

Network interference has attracted significant attention in the field of causal inference, encapsulating various sociological behaviors where the treatment assigned to one individual within a network may affect the outcomes of others, such…

Machine Learning · Computer Science 2025-02-11 Zhiheng Zhang , Zichen Wang

Many real-life optimization problems frequently contain one or more constraints or objectives for which there are no explicit formulas. If data is however available, these data can be used to learn the constraints. The benefits of this…

Machine Learning · Computer Science 2022-09-23 Adejuyigbe Fajemisin , Donato Maragno , Dick den Hertog

In several applications of online optimization to networked systems such as power grids and robotic networks, information about the system model and its disturbances is not generally available. Within the optimization community, increasing…

Optimization and Control · Mathematics 2025-09-01 Caio Kalil Lauand , Andrey Bernstein

Typically, a randomized experiment is designed to test a hypothesis about the average treatment effect and sometimes hypotheses about treatment effect variation. The results of such a study may then be used to inform policy and practice for…

Methodology · Statistics 2026-05-01 Elizabeth Tipton , Michalis Mamakos

Online marketplace designers frequently run A/B tests to measure the impact of proposed product changes. However, given that marketplaces are inherently connected, total average treatment effect estimates obtained through Bernoulli…

Methodology · Statistics 2020-04-28 David Holtz , Ruben Lobel , Inessa Liskovich , Sinan Aral

Practitioners in medicine, business, political science, and other fields are increasingly aware that decisions should be personalized to each patient, customer, or voter. A given treatment (e.g. a drug or advertisement) should be…

Machine Learning · Statistics 2018-06-15 Alejandro Schuler , Michael Baiocchi , Robert Tibshirani , Nigam Shah

Although there is an extensive statistical literature showing the disadvantages of discretizing continuous variables, categorization is a common practice in clinical research which results in substantial loss of information. A large…

Methodology · Statistics 2017-08-17 Márcio Augusto Diniz , Mourad Tighiouart , André Rogatko

It is standard practice in online retail to run pricing experiments by randomizing at the article-level, i.e. by changing prices of different products to identify treatment effects. Due to customers' cross-price substitution behavior, such…

Applications · Statistics 2024-02-23 Lars Roemheld , Justin Rao

Online platforms regularly conduct randomized experiments to understand how changes to the platform causally affect various outcomes of interest. However, experimentation on online platforms has been criticized for having, among other…

Machine Learning · Computer Science 2022-05-12 Smitha Milli , Luca Belli , Moritz Hardt

The estimation of class prevalence, i.e., the fraction of a population that belongs to a certain class, is a very useful tool in data analytics and learning, and finds applications in many domains such as sentiment analysis, epidemiology,…

Machine Learning · Statistics 2021-09-21 Purushottam Kar , Shuai Li , Harikrishna Narasimhan , Sanjay Chawla , Fabrizio Sebastiani

We investigate how to exploit structural similarities of an individual's potential outcomes (POs) under different treatments to obtain better estimates of conditional average treatment effects in finite samples. Especially when it is…

Machine Learning · Statistics 2021-10-26 Alicia Curth , Mihaela van der Schaar

We study hypothesis testing over a heterogeneous population of strategic agents with private information. Any single test applied uniformly across the population yields statistical error that is sub-optimal relative to the performance of an…

Computer Science and Game Theory · Computer Science 2025-10-27 Flora C. Shi , Martin J. Wainwright , Stephen Bates

In this review, we present econometric and statistical methods for analyzing randomized experiments. For basic experiments we stress randomization-based inference as opposed to sampling-based inference. In randomization-based inference,…

Methodology · Statistics 2017-10-26 Susan Athey , Guido Imbens

Model calibration aims to align confidence with prediction correctness. The Cross-Entropy (CE) loss is widely used for calibrator training, which enforces the model to increase confidence on the ground truth class. However, we find the CE…

Computer Vision and Pattern Recognition · Computer Science 2025-02-13 Yuchi Liu , Lei Wang , Yuli Zou , James Zou , Liang Zheng

Assessing the quality of cancer care administered by US health providers poses numerous challenges due to meaningful heterogeneity in patient populations. Patients undergoing oncology treatment exhibit substantial variation in disease…

Applications · Statistics 2025-02-17 Yige Li , Nancy L. Keating , Mary Beth Landrum , Jose R. Zubizarreta

Personalized search is a problem where models benefit from learning user preferences from per-user historical interaction data. The inferred preferences enable personalized ranking models to improve the relevance of documents for users.…

Information Retrieval · Computer Science 2025-05-02 Sheshera Mysore , Garima Dhanania , Kishor Patil , Surya Kallumadi , Andrew McCallum , Hamed Zamani

We study a web-deployed, tool-augmented LLM health coach with real users. In a pilot with seven users (280 rated turns), offline policy evaluation (OPE) over factorized decision heads (Tool/Style) shows that a uniform heavy-tool policy…

Artificial Intelligence · Computer Science 2025-10-22 Melik Ozolcer , Sang Won Bae

Participants in online experiments often enroll over time, which can compromise sample representativeness due to temporal shifts in covariates. This issue is particularly critical in A/B tests, online controlled experiments extensively used…

General Economics · Economics 2026-03-30 Chen Wang , Shichao Han , Shan Huang