中文
相关论文

相关论文: Identifying Heterogeneity in Regression Compositio…

200 篇论文

Income inequality is a major contributor to health disparities, yet its effects often vary by geography and are commonly represented as compositional distributions (e.g., proportions of households across income brackets). Existing spatial…

统计方法学 · 统计学 2026-05-18 Jingwen Deng , Shujie Ma , Sergio J. Rey , Guanyu Hu

Decomposing predictive uncertainty into epistemic (model ignorance) and aleatoric (data ambiguity) components is central to reliable decision making, yet most methods estimate both from the same predictive distribution. Recent empirical and…

机器学习 · 计算机科学 2026-02-13 Tanmoy Mukherjee , Marius Kloft , Pierre Marquis , Zied Bouraoui

High-dimensional data subject to heavy-tailed phenomena and heterogeneity are commonly encountered in various scientific fields and bring new challenges to the classical statistical methods. In this paper, we combine the asymmetric square…

统计理论 · 数学 2019-10-02 Jun Zhao , Guan'ao Yan , Yi Zhang

The rapid growth of high-dimensional datasets across various scientific domains has created a pressing need for new statistical methods to compare distributions supported on their underlying structures. Assessing similarity between datasets…

统计理论 · 数学 2025-11-27 Hongrui Chen , Rong Ma

We address the challenge of solving machine learning tasks using data from privacy-sensitive sellers. Since the data is private, we design a data market that incentivizes sellers to provide their data in exchange for payments. Therefore our…

机器学习 · 计算机科学 2024-10-18 Ameya Anjarlekar , Rasoul Etesami , R. Srikant

Problem definition: Mining for heterogeneous responses to an intervention is a crucial step for data-driven operations, for instance to personalize treatment or pricing. We investigate how to estimate price sensitivity from…

统计方法学 · 统计学 2025-01-08 Jean Pauphilet

Heteroskedastic errors can lead to inaccurate statistical conclusions if they are not properly handled. We introduce a test for heteroskedasticity for the nonparametric regression model with multiple covariates. It is based on a suitable…

统计方法学 · 统计学 2018-02-21 Justin Chown , Ursula U. Müller

Proper scoring rules evaluate the quality of probabilistic predictions, playing an essential role in the pursuit of accurate and well-calibrated models. Every proper score decomposes into two fundamental components -- proper calibration…

Compositional data sets are ubiquitous in science, including geology, ecology, and microbiology. In microbiome research, compositional data primarily arise from high-throughput sequence-based profiling experiments. These data comprise…

统计理论 · 数学 2019-03-05 Patrick L. Combettes , Christian L. Müller

We propose an adjusted likelihood ratio test of two-factor separability (Kronecker product structure) for unbalanced multivariate repeated measures data. Here we address the particular case where the within subject correlation is believed…

统计方法学 · 统计学 2011-02-02 Sean L. Simpson

Compositional data (i.e., data comprising random variables that sum up to a constant) arises in many applications including microbiome studies, chemical ecology, political science, and experimental designs. Yet when compositional data serve…

统计方法学 · 统计学 2025-01-03 Ritwik Bhaduri , Siyuan Ma , Lucas Janson

This paper proposes nonparametric kernel-smoothing estimation for panel data to examine the degree of heterogeneity across cross-sectional units. We first estimate the sample mean, autocovariances, and autocorrelations for each unit and…

计量经济学 · 经济学 2019-05-28 Ryo Okui , Takahide Yanagi

Motivated by regression analysis for microbiome compositional data, this paper considers generalized linear regression analysis with compositional covariates, where a group of linear constraints on regression coefficients are imposed to…

统计方法学 · 统计学 2018-01-11 Jiarui Lu , Pixu Shi , Hongzhe Li

Finding accurate reduced descriptions for large, complex, dynamically evolving networks is a crucial enabler to their simulation, analysis, and, ultimately, design. Here we propose and illustrate a systematic and powerful approach to…

混沌动力学 · 物理学 2017-05-02 Tom Bertalan , Yan Wu , Carlo Laing , C. William Gear , Ioannis G. Kevrekidis

Spatial point pattern data are routinely encountered. A flexible regression model for the underlying intensity is essential to characterizing the spatial point pattern and understanding the impacts of potential risk factors on such pattern.…

统计方法学 · 统计学 2022-12-15 Jieying Jiao , Guanyu Hu , Jun Yan

We propose a method to fuse posterior distributions learned from heterogeneous datasets. Our algorithm relies on a mean field assumption for both the fused model and the individual dataset posteriors and proceeds using a simple…

机器学习 · 计算机科学 2020-07-14 Sebastian Claici , Mikhail Yurochkin , Soumya Ghosh , Justin Solomon

A high-ranking goal of interdisciplinary modeling approaches in the natural sciences are quantitative prediction of system dynamics and model based optimization. For this purpose, mathematical modeling, numerical simulation and scientific…

最优化与控制 · 数学 2015-03-17 Dominik Skanda , Dirk Lebiedz

Estimating Kullback Leibler (KL) divergence from samples of two distributions is essential in many machine learning problems. Variational methods using neural network discriminator have been proposed to achieve this task in a scalable…

机器学习 · 计算机科学 2021-10-01 Sandesh Ghimire , Aria Masoomi , Jennifer Dy

We measure the influence of individual observations on the sequence of the hidden states of the Hidden Markov Model (HMM) by means of the Kullback-Leibler distance (KLD). Namely, we consider the KLD between the conditional distribution of…

信息论 · 计算机科学 2015-06-11 Vittorio Perduca , Gregory Nuel

We consider Heterogeneous Transfer Learning (HTL) from a source to a new target domain for high-dimensional regression with differing feature sets. Most homogeneous TL methods assume that target and source domains share the same feature…

机器学习 · 统计学 2025-12-02 Jae Ho Chang , Massimiliano Russo , Subhadeep Paul