中文
相关论文

相关论文: Bayesian Nonparametric Sensitivity Analysis of Mul…

200 篇论文

Estimating the dependency of variables is a fundamental task in data analysis. Identifying the relevant attributes in databases leads to better data understanding and also improves the performance of learning algorithms, both in terms of…

机器学习 · 计算机科学 2018-10-05 Edouard Fouché , Klemens Böhm

We propose sufficient conditions and computationally efficient procedures for false discovery rate control in multiple testing when the $p$-values are related by a known \emph{dependency graph} -- meaning that we assume independence of…

统计方法学 · 统计学 2025-07-01 Drew T. Nguyen , William Fithian

Epistemic Uncertainty is a measure of the lack of knowledge of a learner which diminishes with more evidence. While existing work focuses on using the variance of the Bayesian posterior due to parameter uncertainty as a measure of epistemic…

In this paper, we introduce a flexible and widely applicable nonparametric entropy-based testing procedure that can be used to assess the validity of simple hypotheses about a specific parametric population distribution. The testing…

计量经济学 · 经济学 2022-01-19 Ron Mittelhammer , George Judge , Miguel Henry

Controlling the false discovery rate (FDR) is a powerful approach to multiple testing. In many applications, the tested hypotheses have an inherent hierarchical structure. In this paper, we focus on the fixed sequence structure where the…

统计方法学 · 统计学 2016-11-11 Gavin Lynch , Wenge Guo , Sanat K. Sarkar , Helmut Finner

Multiple hypothesis testing has been widely applied to problems dealing with high-dimensional data, e.g., selecting significant variables and controlling the selection error rate. The most prevailing measure of error rate used in the…

统计方法学 · 统计学 2022-06-07 Xiaoya Sun , Yan Fu

A sensitivity analysis in an observational study assesses the robustness of significant findings to unmeasured confounding. While sensitivity analyses in matched observational studies have been well addressed when there is a single outcome…

统计方法学 · 统计学 2015-11-05 Colin B. Fogarty , Dylan S. Small

In epidemiology and social sciences, propensity score methods are popular for estimating treatment effects using observational data, and multiple imputation is popular for handling covariate missingness. However, how to appropriately use…

统计方法学 · 统计学 2023-08-30 Trang Quynh Nguyen , Elizabeth A. Stuart

Deep Gaussian Processes (DGPs) were proposed as an expressive Bayesian model capable of a mathematically grounded estimation of uncertainty. The expressivity of DPGs results from not only the compositional character but the distribution…

机器学习 · 计算机科学 2021-11-23 Chi-Ken Lu , Patrick Shafto

This paper presents a survey on some recent advances for the type I error rate control in multiple testing methodology. We consider the problem of controlling the $k$-family-wise error rate (kFWER, probability to make $k$ false discoveries…

统计方法学 · 统计学 2011-03-15 Etienne Roquain

Most causal inference methods consider counterfactual variables under interventions that set the treatment deterministically. With continuous or multi-valued treatments or exposures, such counterfactuals may be of little practical interest…

统计方法学 · 统计学 2021-07-07 Iván Díaz , Nicholas Williams , Katherine L. Hoffman , Edward J. Schenck

The Hierarchical Dirichlet process is a discrete random measure serving as an important prior in Bayesian non-parametrics. It is motivated with the study of groups of clustered data. Each group is modelled through a level two Dirichlet…

概率论 · 数学 2022-10-25 Shui Feng

In many practical applications of multiple hypothesis testing using the False Discovery Rate (FDR), the given hypotheses can be naturally partitioned into groups, and one may not only want to control the number of false discoveries (wrongly…

统计方法学 · 统计学 2016-11-01 Rina Foygel Barber , Aaditya Ramdas

We consider statistical hypothesis testing simultaneously over a fairly general, possibly uncountably infinite, set of null hypotheses, under the assumption that a suitable single test (and corresponding $p$-value) is known for each…

统计方法学 · 统计学 2014-02-10 Gilles Blanchard , Sylvain Delattre , Etienne Roquain

Consider testing multiple hypotheses using tests that can only be evaluated by simulation, such as permutation tests or bootstrap tests. This article introduces MMCTest, a sequential algorithm which gives, with arbitrarily high probability,…

统计方法学 · 统计学 2018-10-17 Axel Gandy , Georg Hahn

Differential privacy (DP) is a rigorous framework that protects the participation of individuals in a dataset by limiting information leakage from released estimators. This creates a challenging setting for statisticians: DP must hold…

统计方法学 · 统计学 2026-05-06 Tao Shen , Xin T. Tong , Wanjie Wang

Multi-fidelity approaches combine different models built on a scarce but accurate data-set (high-fidelity data-set), and a large but approximate one (low-fidelity data-set) in order to improve the prediction accuracy. Gaussian Processes…

机器学习 · 计算机科学 2020-06-30 Ali Hebbal , Loic Brevault , Mathieu Balesdent , El-Ghazali Talbi , Nouredine Melab

Marginal-based methods achieve promising performance in the synthetic data competition hosted by the National Institute of Standards and Technology (NIST). To deal with high-dimensional data, the distribution of synthetic data is…

机器学习 · 计算机科学 2023-01-26 Ximing Li , Chendi Wang , Guang Cheng

This paper explores large sample properties of the two-parameter $(\alpha,\theta)$ Poisson--Dirichlet Process in two contexts. In a Bayesian context of estimating an unknown probability measure, viewing this process as a natural extension…

概率论 · 数学 2008-05-21 Lancelot F. James

In modern scientific experiments, we frequently encounter data that have large dimensions, and in some experiments, such high dimensional data arrive sequentially rather than full data being available all at a time. We develop multiple…

统计方法学 · 统计学 2023-06-09 Rahul Roy , Shyamal K. De , Subir Kumar Bhandari