中文
相关论文

相关论文: Testing the simplifying assumption in high-dimensi…

200 篇论文

The hypothesis that high dimensional data tend to lie in the vicinity of a low dimensional manifold is the basis of manifold learning. The goal of this paper is to develop an algorithm (with accompanying complexity guarantees) for fitting a…

统计理论 · 数学 2013-12-23 Charles Fefferman , Sanjoy Mitter , Hariharan Narayanan

In typical high dimensional statistical inference problems, confidence intervals and hypothesis tests are performed for a low dimensional subset of model parameters under the assumption that the parameters of interest are unconstrained.…

统计方法学 · 统计学 2019-11-19 Ming Yu , Varun Gupta , Mladen Kolar

Uncertain information on input parameters of reliability models is usually modeled by considering these parameters as random, and described by marginal distributions and a dependence structure of these variables. In numerous real-world…

应用统计 · 统计学 2018-04-30 Nazih Benoumechiara , Bertrand Michel , Philippe Saint-Pierre , Nicolas Bousquet

Copulas allow a flexible and simultaneous modeling of complicated dependence structures together with various marginal distributions. Especially if the density function can be represented as the product of the marginal density functions and…

统计方法学 · 统计学 2020-08-31 Jae Youn Ahn , Sebastian Fuchs , Rosy Oh

Simulation offers a simple and flexible way to estimate the power of a clinical trial when analytic formulae are not available. The computational burden of using simulation has, however, restricted its application to only the simplest of…

统计方法学 · 统计学 2020-12-04 Duncan T. Wilson , Rebecca E. A. Walwyn , Richard Hooper , Julia Brown , Amanda J. Farrin

Variable selection plays a fundamental role in high-dimensional data analysis. Various methods have been developed for variable selection in recent years. Well-known examples are forward stepwise regression (FSR) and least angle regression…

统计方法学 · 统计学 2018-02-01 Siliang Gong , Kai Zhang , Yufeng Liu

We investigate an application in the automatic tuning of computer codes, an area of research that has come to prominence alongside the recent rise of distributed scientific processing and heterogeneity in high-performance computing…

应用统计 · 统计学 2013-04-17 Robert B. Gramacy , Matt Taddy , Stefan M. Wild

The problem of testing changes in covariance has received increasing attention in recent years, especially in the context of high-dimensional testing. A number of approaches have been proposed, all limited to the two-sample problem and…

统计方法学 · 统计学 2016-09-06 Yi-Hui Zhou

In this paper, we develop a systematic theory for high dimensional analysis of variance in multivariate linear regression, where the dimension and the number of coefficients can both grow with the sample size. We propose a new \emph{U}~type…

统计方法学 · 统计学 2023-01-12 Zhipeng Lou , Xianyang Zhang , Wei Biao Wu

The ability to adequately model risks is crucial for insurance companies. The method of "Copula-based hierarchical risk aggregation" by Arbenz et al. offers a flexible way in doing so and has attracted much attention recently. We briefly…

风险管理 · 定量金融 2015-06-22 Fabio Derendinger

Intelligent test requires efficient and effective analysis of high-dimensional data in a large scale. Traditionally, the analysis is often conducted by human experts, but it is not scalable in the era of big data. To tackle this challenge,…

机器学习 · 计算机科学 2022-07-04 Yiwen Liao , Tianjie Ge , Raphaël Latty , Bin Yang

The bootstrap is a popular data-driven method to quantify statistical uncertainty, but for modern high-dimensional problems, it could suffer from huge computational costs due to the need to repeatedly generate resamples and refit models. We…

统计方法学 · 统计学 2023-06-21 Henry Lam , Zhenyuan Liu

Estimating copulas with discrete marginal distributions is challenging, especially in high dimensions, because computing the likelihood contribution of each observation requires evaluating $2^{J}$ terms, with $J$ the number of discrete…

统计方法学 · 统计学 2018-11-12 D. Gunawan , M. -N. Tran , K. Suzuki , J. Dick , R. Kohn

We consider the problem of determining feasible systems from a finite set of simulated alternatives with respect to probability constraints, where the observations from stochastic simulations are Bernoulli distributed. Most statistically…

最优化与控制 · 数学 2026-05-27 Taehoon Kim , Sigrun Andradottir , Seong-Hee Kim , Yuwei Zhou

Hypothesis testing in the linear regression model is a fundamental statistical problem. We consider linear regression in the high-dimensional regime where the number of parameters exceeds the number of samples ($p> n$). In order to make…

统计理论 · 数学 2019-09-24 Adel Javanmard , Jason D. Lee

Structured additive distributional copula regression allows to model the joint distribution of multivariate outcomes by relating all distribution parameters to covariates. Estimation via statistical boosting enables accounting for…

This paper studies the problem of high-dimensional multiple testing and sparse recovery from the perspective of sequential analysis. In this setting, the probability of error is a function of the dimension of the problem. A simple…

统计理论 · 数学 2011-06-06 Matthew Malloy , Robert Nowak

Principal component analysis is a versatile tool to reduce dimensionality which has wide applications in statistics and machine learning. It is particularly useful for modeling data in high-dimensional scenarios where the number of…

统计方法学 · 统计学 2022-08-18 Xiaoyu Hu , Fang Yao

High dimensional statistical problems arise from diverse fields of scientific research and technological development. Variable selection plays a pivotal role in contemporary statistical learning and scientific discoveries. The traditional…

统计理论 · 数学 2009-10-08 Jianqing Fan , Jinchi Lv

This paper introduces a novel nonparametric method for estimating high-dimensional dynamic covariance matrices with multiple conditioning covariates, leveraging random forests and supported by robust theoretical guarantees. Unlike…

机器学习 · 统计学 2025-05-20 Shuguang Yu , Fan Zhou , Yingjie Zhang , Ziqi Chen , Hongtu Zhu