中文
相关论文

相关论文: Estimation in exponential family Regression based …

200 篇论文

Heterogeneity is a hallmark of complex diseases. Regression-based heterogeneity analysis, which is directly concerned with outcome-feature relationships, has led to a deeper understanding of disease biology. Such an analysis identifies the…

统计方法学 · 统计学 2022-11-29 Ziye Luo , Xinyue Yao , Yifan Sun , Xinyan Fan

Comparison and contrast are the basic means to unveil causation and learn which treatments work. To build good comparison groups, randomized experimentation is key, yet often infeasible. In such non-experimental settings, we illustrate and…

统计方法学 · 统计学 2024-01-30 Ambarish Chattopadhyay , Jose R. Zubizarreta

Backdoor and data-poisoning attacks can flip predictions with tiny training corruptions, yet a sharp theory linking poisoning strength, overparameterization, and regularization is lacking. We analyze ridge least squares with an unpenalized…

机器学习 · 统计学 2026-01-06 Donald Flynn , Diego Granziol

Missing values in datasets are common in applied statistics. For regression problems, theoretical work thus far has largely considered the issue of missing covariates as distinct from missing responses. However, in practice, many datasets…

统计理论 · 数学 2026-02-17 Benedict M. Risebrow , Thomas B. Berrett

This paper studies sparse linear regression analysis with outliers in the responses. A parameter vector for modeling outliers is added to the standard linear regression model and then the sparse estimation problem for both coefficients and…

统计理论 · 数学 2015-05-21 Shota Katayama , Hironori Fujisawa

This paper investigates correct variable selection in finite samples via $\ell_1$ and $\ell_1+\ell_2$ type penalization schemes. The asymptotic consistency of variable selection immediately follows from this analysis. We focus on logistic…

统计理论 · 数学 2008-12-16 Florentina Bunea

We propose a new class of semiparametric exponential family graphical models for the analysis of high dimensional mixed data. Different from the existing mixed graphical models, we allow the nodewise conditional distributions to be…

机器学习 · 统计学 2015-10-16 Zhuoran Yang , Yang Ning , Han Liu

We propose a new and computationally efficient algorithm for maximizing the observed log-likelihood for a multivariate normal data matrix with missing values. We show that our procedure based on iteratively regressing the missing on the…

统计方法学 · 统计学 2012-11-21 Nicolas Städler , Daniel J. Stekhoven , Peter Bühlmann

By amalgamating data from disparate sources, the resulting integrated dataset becomes a valuable resource for statistical analysis. In probabilistic record linkage, the effectiveness of such integration relies on the availability of linkage…

统计方法学 · 统计学 2025-11-10 Siu-Ming Tam , Min Wang , Alicia Rambaldi , Dehua Tao

Regression calibration as developed by Rosner, Spiegelman and Willet is used to correct the bias in effect estimates due to measurement error in continuous exposures. The method involves two models: a measurement error model (MEM) relating…

统计方法学 · 统计学 2026-02-24 Wenze Tang , Donna Spiegelman , Xiaomei Liao , Molin Wang

We consider generalized linear regression analysis with left-censored covariate due to the lower limit of detection. Complete case analysis by eliminating observations with values below limit of detection yields valid estimates for…

统计方法学 · 统计学 2014-12-09 Shengchun Kong , Bin Nan

In the analysis of data sets consisting of (X, Y)-pairs, a tacit assumption is that each pair corresponds to the same observation unit. If, however, such pairs are obtained via record linkage of two files, this assumption can be violated as…

机器学习 · 统计学 2021-11-03 Zhenbang Wang , Emanuel Ben-David , Martin Slawski

Biomedical researchers usually study the effects of certain exposures on disease risks among a well-defined population. To achieve this goal, the gold standard is to design a trial with an appropriate sample from that population. Due to the…

应用统计 · 统计学 2019-11-18 Cheng Zheng , Sayan Dasgupta , Yuxiang Xie , Asad Haris , Ying Qing Chen

We study a linear observation model with an unknown permutation called \textit{permuted/shuffled linear regression}, where responses and covariates are mismatched and the permutation forms a discrete, factorial-size parameter. The…

统计理论 · 数学 2026-01-23 Hirofumi Ota , Masaaki Imaizumi

Difference-in-differences is a widely-used evaluation strategy that draws causal inference from observational panel data. Its causal identification relies on the assumption of parallel trends, which is scale dependent and may be…

应用统计 · 统计学 2019-06-25 Peng Ding , Fan Li

Computational efficient evaluation of penalized estimators of multivariate exponential family distributions is sought. These distributions encompass among others Markov random fields with variates of mixed type (e.g. binary and continuous)…

统计方法学 · 统计学 2020-12-29 Diederik S. Laman Trip , Wessel N. van Wieringen

Data analysis based on information from several sources is common in economic and biomedical studies. This setting is often referred to as the data fusion problem, which differs from traditional missing data problems since no complete data…

统计方法学 · 统计学 2022-04-07 Wei Li , Shanshan Luo , Wangli Xu

Linear regression and classification methods with repeated functional data are considered. For each statistical unit in the sample, a real-valued parameter is observed over time under different conditions related by some neighborhood…

统计方法学 · 统计学 2024-09-23 Issam-Ali Moindjié , Cristian Preda , Sophie Dabo-Niang

Overlapping asymmetric datasets are common in data science and pose questions of how they can be incorporated together into a predictive analysis. In healthcare datasets there is often a small amount of information that is available for a…

统计方法学 · 统计学 2023-11-21 Matthew McTeer , Robin Henderson , Quentin M Anstee , Paolo Missier

Sparse linear regression -- finding an unknown vector from linear measurements -- is now known to be possible with fewer samples than variables, via methods like the LASSO. We consider the multiple sparse linear regression problem, where…

机器学习 · 计算机科学 2012-02-28 Ali Jalali , Pradeep Ravikumar , Sujay Sanghavi