English
Related papers

Related papers: Decomposition of Differences in Distribution under…

200 papers

The broad objective of this paper is to propose a mathematical model for the study of causes of wage inequality and relate it to choices of consumption, the technologies of production, and the composition of labor in an economy. The paper…

Computer Science and Game Theory · Computer Science 2023-08-21 Sanyukta Deshpande , Milind A. Sohoni

Integrating heterogeneous datasets across different measurement platforms is a fundamental challenge in many scientific applications. A common example arises in deconvolution problems, such as cell type deconvolution, where one aims to…

Methodology · Statistics 2025-09-30 Dongyue Xie , Lin Gui , Jingshu Wang

Machine learning models have achieved widespread success but often inherit and amplify historical biases, resulting in unfair outcomes. Traditional fairness methods typically impose constraints at the prediction level, without addressing…

Machine Learning · Statistics 2026-02-10 Enze Shi , Pankaj Bhagwat , Zhixian Yang , Linglong Kong , Bei Jiang

This paper studies, under the setting of spline regression, the connection between finite-sample properties of selection criteria and their asymptotic counterparts, focusing on bridging the gap between the two. We introduce a bias-variance…

Statistics Theory · Mathematics 2007-06-13 S. C. Kou

Text-to-image models, which can generate high-quality images based on textual input, have recently enabled various content-creation tools. Despite significantly affecting a wide range of downstream applications, the distributions of these…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Yanzhe Zhang , Lu Jiang , Greg Turk , Diyi Yang

This study investigates how personal differences (digital self-efficacy, technical knowledge, belief in equality, political ideology) and demographic factors (age, education, and income) are associated with perceptions of artificial…

Human-Computer Interaction · Computer Science 2024-10-18 Soojong Kim

We propose a novel method for estimating nonseparable selection models. We show that, for a given selection function, the potential outcome distributions are nonparametrically identified from the selected outcome distributions and can be…

Econometrics · Economics 2026-05-05 Fan Wu , Yi Xin

Understanding patterns in mortality across subpopulations is essential for local health policy decision making. One of the key challenges of subnational mortality rate estimation is the presence of small populations and zero or near zero…

Applications · Statistics 2025-12-16 Ameer Dharamshi , Monica Alexander , Celeste Winant , Magali Barbieri

Consider a random sample $X_1 , X_2 , ..., X_n$ drawn independently and identically distributed from some known sampling distribution $P_X$. Let $X_{(1)} \le X_{(2)} \le ... \le X_{(n)}$ represent the order statistics of the sample. The…

Information Theory · Computer Science 2020-09-28 Alex Dytso , Martina Cardone , Cynthia Rush

Inequality measures provide a valuable tool for the analysis, comparison, and optimization based on system models. This work studies the relation between attributes or features of an individual to understand how redundant, unique, and…

Information Theory · Computer Science 2024-07-08 Tobias Mages , Christian Rohner

Probability distributions of money, income, and energy consumption per capita are studied for ensembles of economic agents. The principle of entropy maximization for partitioning of a limited resource gives exponential distributions for the…

Statistical Finance · Quantitative Finance 2010-09-02 Anand Banerjee , Victor M. Yakovenko

This paper introduces a machine for sampling approximate model-X knockoffs for arbitrary and unspecified data distributions using deep generative models. The main idea is to iteratively refine a knockoff sampling mechanism until a criterion…

Methodology · Statistics 2020-03-03 Yaniv Romano , Matteo Sesia , Emmanuel J. Candès

Feature selection can facilitate the learning of mixtures of discrete random variables as they arise, e.g. in crowdsourcing tasks. Intuitively, not all workers are equally reliable but, if the less reliable ones could be eliminated, then…

Machine Learning · Statistics 2017-11-28 Vincent Zhao , Steven W. Zucker

Causal decomposition has provided a powerful tool to analyze health disparity problems, by assessing the proportion of disparity caused by each mediator. However, most of these methods lack \emph{policy implications}, as they fail to…

Methodology · Statistics 2023-02-21 Xinwei Sun , Xiangyu Zheng , Jim Weinstein

Understanding citizens' values in participatory systems is crucial for citizen-centric policy-making. We envision a hybrid participatory system where participants make choices and provide motivations for those choices, and AI agents…

Artificial Intelligence · Computer Science 2025-02-12 Enrico Liscio , Luciano C. Siebert , Catholijn M. Jonker , Pradeep K. Murukannaiah

With the rise of increasingly powerful and user-facing NLP systems, there is growing interest in assessing whether they have a good representation of uncertainty by evaluating the quality of their predictive distribution over outcomes. We…

Computation and Language · Computer Science 2024-02-27 Joris Baan , Raquel Fernández , Barbara Plank , Wilker Aziz

This article investigates the fundamental factors influencing the rate and manner of Electoral participation with an economic model-based approach. In this study, the structural parameters affecting people's decision making are divided into…

General Economics · Economics 2025-10-29 Mostafa Raeisi Sarkandiz

We discuss the distribution of commuting distances and its relation to income. Using data from Denmark, the UK, and the US, we show that the commuting distance is (i) broadly distributed with a slow decaying tail that can be fitted by a…

Physics and Society · Physics 2016-06-13 Giulia Carra , Ismir Mulalic , Mogens Fosgerau , Marc Barthelemy

In the linear mixed model (LMM), the simultaneous assessment and comparison of dispersion relevance of explanatory variables associated with fixed and random effects remains an important open practical problem. Based on the restricted…

Methodology · Statistics 2023-05-31 Nicholas Schreck , Manuel Wiesenfarth

We propose autoregressive Bayesian semi-parametric models for waiting times between recurrent events. The aim is two-fold: inference on the effect of possibly time-varying covariates on the gap times and clustering of individuals based on…

Applications · Statistics 2016-07-28 Marta Tallarita , Maria De Iorio , Alessandra Guglielmi , James Malone-Lee