中文
相关论文

相关论文: A stable and adaptive polygenic signal detection m…

200 篇论文

Biased sampling designs can be highly efficient when studying rare (binary) or low variability (continuous) endpoints. We consider longitudinal data settings in which the probability of being sampled depends on a repeatedly measured…

统计方法学 · 统计学 2020-01-14 Lee S. McDaniel , Jonathan S. Schildcrout , Enrique F. Schisterman , Paul J. Rathouz

Variable selection for high-dimensional, highly correlated data has long been a challenging problem, often yielding unstable and unreliable models. We propose a resample-aggregate framework that exploits diffusion models' ability to…

统计方法学 · 统计学 2025-08-20 Minjie Wang , Xiaotong Shen , Wei Pan

We propose a general, modular method for significance testing of groups (or clusters) of variables in a high-dimensional linear model. In presence of high correlations among the covariables, due to serious problems of identifiability, it is…

统计理论 · 数学 2015-02-12 Jacopo Mandozzi , Peter Bühlmann

Conventional multiple testing procedures often assume hypotheses for different features are exchangeable. However, in many scientific applications, additional covariate information regarding the patterns of signals and nulls are available.…

统计方法学 · 统计学 2020-06-12 Xianyang Zhang , Jun Chen

This paper introduces a new method for change detection in psychometric studies based on the recently introduced pseudo Score statistic, for which the sampling distribution under the alternative hypothesis has been determined. Our approach…

统计方法学 · 统计学 2024-08-09 Nicoletta D'Angelo

In this paper, we consider the problem of testing properties of joint distributions under the Conditional Sampling framework. In the standard sampling model, the sample complexity of testing properties of joint distributions is exponential…

计算复杂性 · 计算机科学 2022-08-03 Rishiraj Bhattacharyya , Sourav Chakraborty

This paper introduces a stochastic plug-and-play (PnP) sampling algorithm that leverages variable splitting to efficiently sample from a posterior distribution. The algorithm based on split Gibbs sampling (SGS) draws inspiration from the…

机器学习 · 统计学 2023-04-24 Florentin Coeurdoux , Nicolas Dobigeon , Pierre Chainais

In unsupervised learning, dimensionality reduction is an important tool for data exploration and visualization. Because these aims are typically open-ended, it can be useful to frame the problem as looking for patterns that are enriched in…

机器学习 · 统计学 2018-11-16 Kristen Severson , Soumya Ghosh , Kenney Ng

The detection of a signal variable from multiple variables that contain many noise variables is often approached as a variable selection problem under a given objective variable. This is nothing more than building a supervised model of a…

数据分析、统计与概率 · 物理学 2025-09-29 Yoh-ichi Mototake , Y-h. Taguchi

This work is devoted to the development of a distributionally robust active fault diagnosis approach for a class of nonlinear systems, which takes into account any ambiguity in distribution information of the uncertain model parameters.…

最优化与控制 · 数学 2021-08-12 Ioannis Tzortzis , Marios M. Polycarpou

For data segmentation in high-dimensional linear regression settings, the regression parameters are often assumed to be sparse segment-wise, which enables many existing methods to estimate the parameters locally via $\ell_1$-regularised…

统计方法学 · 统计学 2026-05-08 Haeran Cho , Tobias Kley , Housen Li

We present the extention and application of a new unsupervised statistical learning technique--the Partition Decoupling Method--to gene expression data. Because it has the ability to reveal non-linear and non-convex geometries present in…

定量方法 · 定量生物学 2015-09-24 Rosemary Braun , Gregory Leibon , Scott Pauls , Daniel Rockmore

The parallel alternating direction method of multipliers (ADMM) algorithms have gained popularity in statistics and machine learning due to their efficient handling of large sample data problems. However, the parallel structure of these…

统计理论 · 数学 2024-04-11 Xiaofei Wu , Jiancheng Jiang , Zhimin Zhang

Approaches for testing sets of variants, such as a set of rare or common variants within a gene or pathway, for association with complex traits are important. In particular, set tests allow for aggregation of weak signal within a set, can…

基因组学 · 定量生物学 2013-05-28 Jennifer Listgarten , Christoph Lippert , Eun Yong Kang , Jing Xiang , Carl M. Kadie , David Heckerman

This work is concerned with the detection of a mixture distribution from a $\mathbb{R}$-valued sample. Given a sample $X_1,\dots,X_n$ and an even density $\phi$, our aim is to detect whether the sample distribution is $\phi(\cdot-\mu)$ for…

统计理论 · 数学 2016-01-22 Béatrice Laurent , Clément Marteau , Cathy Maugis-Rabusseau

We introduce a new method of performing high dimensional discriminant analysis, which we call multiDA. We achieve this by constructing a hybrid model that seamlessly integrates a multiclass diagonal discriminant analysis model and feature…

机器学习 · 统计学 2018-07-05 Sarah Elizabeth Romanes , John Thomas Ormerod , Jean YH Yang

A signal recovery problem is considered, where the same binary testing problem is posed over multiple, independent data streams. The goal is to identify all signals, i.e., streams where the alternative hypothesis is correct, and noises,…

统计方法学 · 统计学 2022-11-08 Yiming Xing , Georgios Fellouris

Matched case-control studies are commonly employed in epidemiological research for their convenience and efficiency. Analysis of secondary outcomes can yield valuable insights into biological pathways and help identify genetic variants of…

统计方法学 · 统计学 2026-02-24 Shanshan Liu , Guoqing Diao

Stable distributions provide a flexible framework for modeling heavy-tailed and skewed data, with the stability index $\alpha$ quantifying tail heaviness. We propose a new semiparametric estimator for $\alpha$ that leverages the two-sum…

统计方法学 · 统计学 2025-08-19 Cornelis J. Potgieter , Jacques van Appel , Sudharshan Samaratunga

Unsupervised feature selection is an important method to reduce dimensions of high dimensional data without labels, which is benefit to avoid ``curse of dimensionality'' and improve the performance of subsequent machine learning tasks, like…

机器学习 · 计算机科学 2020-12-29 Yanyong Huang , Zongxin Shen , Fuxu Cai , Tianrui Li , Fengmao Lv