中文
相关论文

相关论文: On valid descriptive inference from non-probabilit…

200 篇论文

Statistical modeling plays a fundamental role in understanding the underlying mechanism of massive data (statistical inference) and predicting the future (statistical prediction). Although all models are wrong, researchers try their best to…

统计方法学 · 统计学 2020-06-17 Hangjin Jiang

What is the difference of a prediction that is made with a causal model and a non-causal model? Suppose we intervene on the predictor variables or change the whole environment. The predictions from a causal model will in general work as…

统计方法学 · 统计学 2024-04-27 Jonas Peters , Peter Bühlmann , Nicolai Meinshausen

In order to trust the predictions of a machine learning algorithm, it is necessary to understand the factors that contribute to those predictions. In the case of probabilistic and uncertainty-aware models, it is necessary to understand not…

机器学习 · 统计学 2024-08-19 Danny Wood , Theodore Papamarkou , Matt Benatan , Richard Allmendinger

Cross-validation is a popular non-parametric method for evaluating the accuracy of a predictive rule. The usefulness of cross-validation depends on the task we want to employ it for. In this note, I discuss a simple non-parametric setting,…

统计方法学 · 统计学 2019-09-27 Stefan Wager

We consider the problem of distribution-free predictive inference, with the goal of producing predictive coverage guarantees that hold conditionally rather than marginally. Existing methods such as conformal prediction offer marginal…

We consider the problem of subspace estimation in situations where the number of available snapshots and the observation dimension are comparable in magnitude. In this context, traditional subspace methods tend to fail because the…

信息论 · 计算机科学 2016-11-15 Pascal Vallet , Philippe Loubaton , Xavier Mestre

In this paper, we develop invariance-based procedures for testing and inference in high-dimensional regression models. These procedures, also known as randomization tests, provide several important advantages. First, for the global null…

统计方法学 · 统计学 2023-12-27 Wenxuan Guo , Panos Toulis

Generalization to new samples is a fundamental rationale for statistical modeling. For this purpose, model validation is particularly important, but recent work in survey inference has suggested that simple aggregation of individual…

统计方法学 · 统计学 2024-04-15 Lauren Kennedy , Aki Vehtari , Andrew Gelman

For obtaining causal inferences that are objective, and therefore have the best chance of revealing scientific truths, carefully designed and executed randomized experiments are generally considered to be the gold standard. Observational…

应用统计 · 统计学 2008-11-12 Donald B. Rubin

In this paper, we address the probabilistic error quantification of a general class of prediction methods. We consider a given prediction model and show how to obtain, through a sample-based approach, a probabilistic upper bound on the…

统计理论 · 数学 2021-06-07 Victor Mirasierra , Martina Mammarella , Fabrizio Dabbene , Teodoro Alamo

Causal inference from observational data often assumes "ignorability," that all confounders are observed. This assumption is standard yet untestable. However, many scientific studies involve multiple causes, different variables whose…

机器学习 · 统计学 2019-04-16 Yixin Wang , David M. Blei

We tackle the problem of conditioning probabilistic programs on distributions of observable variables. Probabilistic programs are usually conditioned on samples from the joint data distribution, which we refer to as deterministic…

机器学习 · 计算机科学 2021-03-09 David Tolpin , Yuan Zhou , Tom Rainforth , Hongseok Yang

For binary experimental data, we discuss randomization-based inferential procedures that do not need to invoke any modeling assumptions. We also introduce methods for likelihood and Bayesian inference based solely on the physical…

统计方法学 · 统计学 2017-05-25 Peng Ding , Luke W. Miratrix

This paper is devoted to establishing exponential bounds for the probabilities of deviation of a sample sum from its expectation, when the variables involved in the summation are obtained by sampling in a finite population according to a…

统计理论 · 数学 2016-10-13 Patrice Bertail , Stephan Clémençon

Selective classification allows models to abstain from making predictions (e.g., say "I don't know") when in doubt in order to obtain better effective accuracy. While typical selective models can be effective at producing more accurate…

机器学习 · 计算机科学 2024-06-24 Adam Fisch , Tommi Jaakkola , Regina Barzilay

We consider the problem of inference after model selection under weak assumptions in the time series setting. Even when the data are not independent, we show that sample splitting remains asymptotically valid as long as the process…

统计理论 · 数学 2019-02-27 Robert Lunde

We explore the interplay between random and deterministic phenomena using a representation of uncertainty based on the measure-theoretic concept of outer measure. The meaning of the analogues of different probabilistic concepts is…

统计方法学 · 统计学 2020-04-21 Jeremie Houssineau

Inference in current domains of application are often complex and require us to integrate the expertise of a variety of disparate panels of experts and models coherently. In this paper we develop a formal statistical methodology to guide…

统计方法学 · 统计学 2018-07-30 Manuele Leonelli , Martine J. Barons , Jim Q. Smith

We study the assessment of semiparametric and other highly-parametrised models from the perspective of foundational principles of parametric statistical inference. In doing so, we highlight the possibility of avoiding the usual…

统计方法学 · 统计学 2026-05-05 Heather Battey , Nancy Reid

After variable selection, standard inferential procedures for regression parameters may not be uniformly valid; there is no finite-sample size at which a standard test is guaranteed to approximately attain its nominal size. This problem is…

统计方法学 · 统计学 2020-07-07 Oliver Dukes , Vahe Avagyan , Stijn Vansteelandt