中文
相关论文

相关论文: High-Dimensional Inference: Confidence Intervals, …

200 篇论文

While there is substantial need for dependence models in higher dimensions, most existing models quickly become rather restrictive and barely balance parsimony and flexibility. Hierarchical constructions may improve on that by grouping…

统计方法学 · 统计学 2013-10-11 Eike Christian Brechmann

Although a majority of the theoretical literature in high-dimensional statistics has focused on settings which involve fully-observed data, settings with missing values and corruptions are common in practice. We consider the problems of…

机器学习 · 统计学 2017-11-06 Yining Wang , Jialei Wang , Sivaraman Balakrishnan , Aarti Singh

When analyzing empirical data, we often find that global linear models overestimate the number of parameters required. In such cases, we may ask whether the data lies on or near a manifold or a set of manifolds (a so-called multi-manifold)…

机器学习 · 统计学 2018-07-03 F. Patricia Medina , Linda Ness , Melanie Weber , Karamatou Yacoubou Djima

By seeking the narrowest prediction intervals (PIs) that satisfy the specified coverage probability requirements, the recently proposed quality-based PI learning principle can extract high-quality PIs that better summarize the predictive…

机器学习 · 计算机科学 2019-07-23 Lin Zhu , Jiaxing Lu , Yihong Chen

Prediction algorithms, such as deep neural networks (DNNs), are used in many domain sciences to directly estimate internal parameters of interest in simulator-based models, especially in settings where the observations include images or…

机器学习 · 统计学 2023-11-14 Luca Masserano , Tommaso Dorigo , Rafael Izbicki , Mikael Kuusela , Ann B. Lee

It is critical to accurately simulate data when employing Monte Carlo techniques and evaluating statistical methodology. Measurements are often correlated and high dimensional in this era of big data, such as data obtained in…

We present a Python implementation for RS-HDMR-GPR (Random Sampling High Dimensional Model Representation Gaussian Process Regression). The method builds representations of multivariate functions with lower-dimensional terms, either as an…

统计计算 · 统计学 2023-01-27 Owen Ren , Mohamed Ali Boussaidi , Dmitry Voytsekhovsky , Manabu Ihara , Sergei Manzhos

Graphical models have become a very popular tool for representing dependencies within a large set of variables and are key for representing causal structures. We provide results for uniform inference on high-dimensional graphical models…

统计方法学 · 统计学 2018-12-04 Sven Klaassen , Jannis Kück , Martin Spindler , Victor Chernozhukov

This article considers a novel and widely applicable approach to modeling high-dimensional dependent data when a large number of explanatory variables are available and the signal-to-noise ratio is low. We postulate that a $p$-dimensional…

统计方法学 · 统计学 2024-12-09 Zhaoxing Gao , Ruey S. Tsay

Analyzing time series in the frequency domain enables the development of powerful tools for investigating the second-order characteristics of multivariate processes. Parameters like the spectral density matrix and its inverse, the coherence…

统计方法学 · 统计学 2024-01-19 Jonas Krampe , Efstathios Paparoditis

We derive high-dimensional Gaussian comparison results for the standard $V$-fold cross-validated risk estimates. Our results combine a recent stability-based argument for the low-dimensional central limit theorem of cross-validation with…

统计理论 · 数学 2023-11-15 Nicholas Kissel , Jing Lei

Latent factor models that integrate data from multiple sources/studies or modalities have garnered considerable attention across various disciplines. However, existing methods predominantly focus either on multi-study integration or…

统计方法学 · 统计学 2025-07-15 Wei Liu , Qingzhi Zhong

It is clear that conventional statistical inference protocols need to be revised to deal correctly with the high-dimensional data that are now common. Most recent studies aimed at achieving this revision rely on powerful approximation…

We consider variable selection in high-dimensional linear models where the number of covariates greatly exceeds the sample size. We introduce the new concept of partial faithfulness and use it to infer associations between the covariates…

统计方法学 · 统计学 2012-01-12 Peter Bühlmann , Markus Kalisch , Marloes H. Maathuis

The age of big data has produced data sets that are computationally expensive to analyze and store. Algorithmic leveraging proposes that we sample observations from the original data set to generate a representative data set and then…

应用统计 · 统计学 2018-03-13 Katelyn Gao

Summary: ipd is an open-source R software package for the downstream modeling of an outcome and its associated features where a potentially sizable portion of the outcome data has been imputed by an artificial intelligence or machine…

Regression analysis of correlated data, where multiple correlated responses are recorded on the same unit, is ubiquitous in many scientific areas. With the advent of new technologies, in particular high-throughput omics profiling assays,…

统计方法学 · 统计学 2024-09-04 Lu Xia , Ali Shojaie

Linear mixed models (LMMs) are used extensively to model dependecies of observations in linear regression and are used extensively in many application areas. Parameter estimation for LMMs can be computationally prohibitive on big data.…

机器学习 · 统计学 2019-03-08 Zilong Tan , Kimberly Roche , Xiang Zhou , Sayan Mukherjee

I propose a new type of confidence interval for correct asymptotic inference after using data to select a model of interest without assuming any model is correctly specified. This hybrid confidence interval is constructed by combining…

统计方法学 · 统计学 2021-11-25 Adam McCloskey

High-dimensional penalized rank regression is a powerful tool for modeling high-dimensional data due to its robustness and estimation efficiency. However, the non-smoothness of the rank loss brings great challenges to the computation. To…

统计方法学 · 统计学 2025-02-20 Leheng Cai , Xu Guo , Heng Lian , Liping Zhu