English
Related papers

Related papers: Proxy expenditure weights for Consumer Price Index…

200 papers

Confounding remains one of the major challenges to causal inference with observational data. This problem is paramount in medicine, where we would like to answer causal questions from large observational datasets like electronic health…

Methodology · Statistics 2024-01-10 Linying Zhang , Yixin Wang , Martijn Schuemie , David Blei , George Hripcsak

Online data has the potential to transform how researchers and companies produce election forecasts. Social media surveys, online panels and even comments scraped from the internet can offer valuable insights into political preferences.…

Applications · Statistics 2025-03-20 Alberto Arletti , Maria Letizia Tanturri , Omar Paccagnella

A long-standing question about consumer behavior is whether individuals' observed purchase decisions satisfy the revealed preference (RP) axioms of the utility maximization theory (UMT). Researchers using survey or experimental panel data…

Econometrics · Economics 2020-09-21 Victor H. Aguiar , Nail Kashaev

Neural network approaches in recommender systems have shown remarkable success by representing a large set of items as a learnable vector embedding table. However, infrequent items may suffer from inadequate training opportunities, making…

Information Retrieval · Computer Science 2023-12-12 Jinseok Seol , Minseok Gang , Sang-goo Lee , Jaehui Park

Survey sampling is concerned with the estimation of finite population parameters. In practice, survey data suffer from item nonresponse, which is commonly handled through imputation, i.e., replacing missing values with predicted values. As…

Methodology · Statistics 2026-03-06 Ziming An , Mehdi Dagdoug , David Haziza

Data integration has become increasingly popular owing to the availability of multiple data sources. This study considered quantile regression estimation when a key covariate had multiple proxies across several datasets. In a unified…

Methodology · Statistics 2022-10-25 Dongyoung Go , Jongho Im , Ick Hoon Jin

We consider a situation where the distribution of a random variable is being estimated by the empirical distribution of noisy measurements of that variable. This is common practice in, for example, teacher value-added models and other…

Econometrics · Economics 2021-12-08 Koen Jochmans , Martin Weidner

Due to their heterogeneity, insurance risks can be properly described as a mixture of different fixed models, where the weights assigned to each model may be estimated empirically from a sample of available data. If a risk measure is…

Risk Management · Quantitative Finance 2018-02-12 Valeria Bignozzi , Claudio Macci , Lea Petrella

We develop several tools for the determination of sample size and design for MediCal audits. This audit setting involves a population of claims for reimbursement by a healthcare provider which need to be reviewed by an auditor to determine…

Methodology · Statistics 2018-02-13 Michelle Norris

We study statistical parameter estimation in the setting of data markets. A buyer seeks to estimate a parameter based on samples that can be purchased from competing providers that differ in their data quality and provision costs. When…

Computer Science and Game Theory · Computer Science 2026-04-13 Yuchen Hu , Martin J. Wainwright , Stephen Bates

Sensitivity analysis is widely used to assess the robustness of causal conclusions in observational studies, yet its interaction with the structure of measured covariates is often overlooked. When latent confounders cannot be directly…

Methodology · Statistics 2026-02-17 Abhinandan Dalal , Iris Horng , Yang Feng , Dylan S. Small

The US Census Bureau will deliberately corrupt data sets derived from the 2020 US Census, enhancing the privacy of respondents while potentially reducing the precision of economic analysis. To investigate whether this trade-off is…

Econometrics · Economics 2024-02-13 Anish Agarwal , Rahul Singh

Mainstream bias, where some users receive poor recommendations because their preferences are uncommon or simply because they are less active, is an important aspect to consider regarding fairness in recommender systems. Existing methods to…

Information Retrieval · Computer Science 2023-07-26 Roger Zhe Li , Julián Urbano , Alan Hanjalic

Intertemporal choices involve making decisions that require weighing the costs in the present against the benefits in the future. One specific type of intertemporal choice is the decision between purchasing an individual item or opting for…

Information Retrieval · Computer Science 2023-09-20 Qingming Li , H. Vicky Zhao

Measurement error in the covariate of main interest (e.g. the exposure variable, or the risk factor) is common in epidemiologic and health studies. It can effect the relative risk estimator or other types of coefficients derived from the…

Methodology · Statistics 2022-07-20 Sarit Agami

Survey data typically have missing values due to unit and item nonresponse. Sometimes, survey organizations know the marginal distributions of certain categorical variables in the survey. As shown in previous work, survey organizations can…

Methodology · Statistics 2025-09-01 Kewei Xu , Jerome P. Reiter

We study a data analyst's problem of acquiring data from self-interested individuals to obtain an accurate estimation of some statistic of a population, subject to an expected budget constraint. Each data holder incurs a cost, which is…

Computer Science and Game Theory · Computer Science 2019-05-15 Yiling Chen , Shuran Zheng

The importance of exploring a potential integration among surveys has been acknowledged in order to enhance effectiveness and minimize expenses. In this work, we employ the alignment method to combine information from two different surveys…

Methodology · Statistics 2024-04-09 Vasilis Chasiotis , Dimitris Karlis

Reliance on images for dietary assessment is an important strategy to accurately and conveniently monitor an individual's health, making it a vital mechanism in the prevention and care of chronic diseases and obesity. However, image-based…

Computer Vision and Pattern Recognition · Computer Science 2026-02-06 Gautham Vinod , Fengqing Zhu

In machine learning, a bias occurs whenever training sets are not representative for the test data, which results in unreliable models. The most common biases in data are arguably class imbalance and covariate shift. In this work, we aim to…

Machine Learning · Computer Science 2018-04-04 Patrick Glauner , Radu State , Petko Valtchev , Diogo Duarte