中文
相关论文

相关论文: Efficient Multiple Testing Adjustment for Hierarch…

200 篇论文

After initial release of a machine learning algorithm, the model can be fine-tuned by retraining on subsequently gathered data, adding newly discovered features, or more. Each modification introduces a risk of deteriorating performance and…

Usually one compares the accuracy of two competing classifiers via null hypothesis significance tests (nhst). Yet the nhst tests suffer from important shortcomings, which can be overcome by switching to Bayesian hypothesis testing. We…

机器学习 · 计算机科学 2016-11-23 Giorgio Corani , Alessio Benavoli , Janez Demšar , Francesca Mangili , Marco Zaffalon

This paper develops a large-scale inference approach for the regularization of stock return covariance matrices. The framework allows for the presence of heavy tails and multivariate GARCH-type effects of unknown form among the stock…

计量经济学 · 经济学 2024-07-16 Richard Luger

The level set approach has proven widely successful in the study of inverse problems for interfaces, since its systematic development in the 1990s. Recently it has been employed in the context of Bayesian inversion, allowing for the…

概率论 · 数学 2016-09-13 Matthew M. Dunlop , Marco A. Iglesias , Andrew M. Stuart

In this paper, we develop a systematic theory for high dimensional analysis of variance in multivariate linear regression, where the dimension and the number of coefficients can both grow with the sample size. We propose a new \emph{U}~type…

统计方法学 · 统计学 2023-01-12 Zhipeng Lou , Xianyang Zhang , Wei Biao Wu

This work introduces an adaptive mesh refinement technique for hierarchical hybrid grids with the goal to reach scalability and maintain excellent performance on massively parallel computer systems. On the block structured hierarchical…

数值分析 · 数学 2025-08-11 Benjamin Mann , Ulrich Rüde

Variational families with full-rank covariance approximations are known not to work well in black-box variational inference (BBVI), both empirically and theoretically. In fact, recent computational complexity results for BBVI have…

机器学习 · 统计学 2025-11-14 Joohwan Ko , Kyurae Kim , Woo Chang Kim , Jacob R. Gardner

This paper develops a general framework for conducting inference on the rank of an unknown matrix $\Pi_0$. A defining feature of our setup is the null hypothesis of the form $\mathrm H_0: \mathrm{rank}(\Pi_0)\le r$. The problem is of first…

计量经济学 · 经济学 2019-03-26 Qihui Chen , Zheng Fang

Hierarchical model fitting has become commonplace for case-control studies of cognition and behaviour in mental health. However, these techniques require us to formalise assumptions about the data-generating process at the group level,…

计算机与社会 · 计算机科学 2020-11-04 Vincent Valton , Toby Wise , Oliver J. Robinson

High dimensional hypothesis test deals with models in which the number of parameters is significantly larger than the sample size. Existing literature develops a variety of individual tests. Some of them are sensitive to the dense and small…

统计理论 · 数学 2018-08-09 Cheng Zhou , Xinsheng Zhang , Wenxin Zhou , Han Liu

In modern multilabel classification problems, each data instance belongs to a small number of classes from a large set of classes. In other words, these problems involve learning very sparse binary label vectors. Moreover, in large-scale…

机器学习 · 计算机科学 2020-11-03 Shashanka Ubaru , Sanjeeb Dash , Arya Mazumdar , Oktay Gunluk

We consider inference in linear regression models that is robust to heteroskedasticity and the presence of many control variables. When the number of control variables increases at the same rate as the sample size the usual…

统计理论 · 数学 2020-09-29 Koen Jochmans

High-dimensional tests are applied to find relevant sets of variables and relevant models. If variables are selected by analyzing the sums of products matrices and a corresponding mean-value test is performed, there is the danger that the…

统计方法学 · 统计学 2012-02-10 Juergen Laeuter , Maciej Rosolowski , Ekkehard Glimm

Biomarker subpopulations have become increasingly important for drug development in targeted therapies. The use of biomarkers has the potential to facilitate more effective outcomes by guiding patient selection appropriately, thus enhancing…

统计方法学 · 统计学 2020-08-07 Ting-Yu Chen , Jing Zhao , Linda Sun , Keaven Anderson

An important task for any large-scale organization is to prepare forecasts of key performance metrics. Often these organizations are structured in a hierarchical manner and for operational reasons, projections of these metrics may have been…

应用统计 · 统计学 2017-11-15 Julie Novak , Scott McGarvie , Beatriz Etchegaray Garcia

This work introduces a novel, simple, and flexible method to quantify irreversibility in generic high-dimensional time series based on the well-known mapping to a binary classification problem. Our approach utilizes gradient boosting for…

统计力学 · 物理学 2025-01-09 Michele Vodret , Cristiano Pacini , Christian Bongiorno

Modern high-throughput biomedical devices routinely produce data on a large scale, and the analysis of high-dimensional datasets has become commonplace in biomedical studies. However, given thousands or tens of thousands of measured…

统计方法学 · 统计学 2022-02-28 Vladimir Vutov , Thorsten Dickhaus

Estimating a covariance matrix is central to high-dimensional data analysis. Empirical analyses of high-dimensional biomedical data, including genomics, proteomics, microbiome, and neuroimaging, among others, consistently reveal strong…

统计方法学 · 统计学 2024-12-05 Yifan Yang , Chixiang Chen , Shuo Chen

We consider the problem of selecting confounders for adjustment from a potentially large set of covariates, when estimating a causal effect. Recently, the high-dimensional Propensity Score (hdPS) method was developed for this task; hdPS…

统计方法学 · 统计学 2021-12-17 Asad Haris , Robert Platt

We propose a new method for modelling simple longitudinal data. We aim to do this in a flexible manner (without restrictive assumptions about the shapes of individual trajectories), while exploiting structural similarities between the…

统计方法学 · 统计学 2024-09-24 Helen Ogden