English
Related papers

Related papers: The Hiemstra-Jones Test Revisited

200 papers

Propensity score methods have become a part of the standard toolkit for applied researchers who wish to ascertain causal effects from observational data. While they were originally developed for binary treatments, several researchers have…

Methodology · Statistics 2013-09-26 Shandong Zhao , David A. van Dyk , Kosuke Imai

This is an up-to-date introduction to, and overview of, marginal likelihood computation for model selection and hypothesis testing. Computing normalizing constants of probability models (or ratio of constants) is a fundamental issue in many…

Computation · Statistics 2023-02-13 Fernando Llorente , Luca Martino , David Delgado , Javier Lopez-Santiago

We study estimation and inference for heterogeneous principal causal effects with binary treatments and binary intermediate variables. Principal causal effects are subgroup effects within strata defined by potential values of an…

Methodology · Statistics 2026-03-11 Rui Zhang , Charles R. Doss , Jared D. Huling

This paper proposes a new approach for testing Granger non-causality on panel data. Instead of aggregating panel member statistics, we aggregate their corresponding p-values and show that the resulting p-value approximately bounds the type…

The aim of this paper is to establish non-asymptotic minimax rates of testing for goodness-of-fit hypotheses in a heteroscedastic setting. More precisely, we deal with sequences $(Y_j)_{j\in J}$ of independent Gaussian random variables,…

Statistics Theory · Mathematics 2010-02-09 Béatrice Laurent , Jean-Michel Loubès , Clément Marteau

External validity is often questionable in empirical research, especially in randomized experiments due to the trade-off between internal validity and external validity. To quantify the robustness of external validity, one must first…

Methodology · Statistics 2022-06-20 Tenglong Li

In this paper, we develop invariance-based procedures for testing and inference in high-dimensional regression models. These procedures, also known as randomization tests, provide several important advantages. First, for the global null…

Methodology · Statistics 2023-12-27 Wenxuan Guo , Panos Toulis

I proposed (8, 1, 3) that p values should be supplemented by an estimate of the false positive risk (FPR). FPR was defined as the probability that, if you claim that there is a real effect on the basis of p value from a single unbiased…

Other Statistics · Statistics 2020-08-10 David Colquhoun

Evaluating natural language generation (NLG) systems remains a core challenge of natural language processing (NLP), further complicated by the rise of large language models (LLMs) that aims to be general-purpose. Recently, large language…

Computation and Language · Computer Science 2025-08-29 Khaoula Chehbouni , Mohammed Haddou , Jackie Chi Kit Cheung , Golnoosh Farnadi

In this article, we propose a new class of consistent tests for $p$-variate normality. These tests are based on the characterization of the standard multivariate normal distribution, that the Hessian of the corresponding cumulant generating…

Methodology · Statistics 2023-03-22 Kwun Chuen Gary Chan , Hok Kan Ling , Chuan-Fa Tang , Sheung Chi Phillip Yam

Long-term causal inference has drawn increasing attention in many scientific domains. Existing methods mainly focus on estimating average long-term causal effects by combining long-term observational data and short-term experimental data.…

Machine Learning · Computer Science 2025-03-04 Weilin Chen , Ruichu Cai , Junjie Wan , Zeqin Yang , José Miguel Hernández-Lobato

Nonresponse weighting adjustment using the response propensity score is a popular tool for handling unit nonresponse. Statistical inference after the nonresponse weighting adjustment is complicated because the effect of estimating the…

Methodology · Statistics 2017-02-14 Hejian Sang , Jae Kwang Kim

Gibbs-type random probability measures and the exchangeable random partitions they induce represent the subject of a rich and active literature. They provide a probabilistic framework for a wide range of theoretical and applied problems…

Statistics Theory · Mathematics 2015-04-06 Sergio Bacallado , Stefano Favaro , Lorenzo Trippa

In this paper it is shown that in the case of the given probability distribution or histogram using the Jaynes entropy maximum principle an J- measure of uncertainty (JMU) provides the maximum for a given distribution can be constructed.…

Statistical Mechanics · Physics 2019-05-29 Y. Schreiber , A. Chudnovsky

Let $X_1,X_2, \ldots$ be independent and identically distributed random elements taking values in a separable Hilbert space $\mathbb{H}$. With applications for functional data in mind, $\mathbb{H}$ may be regarded as a space of…

Statistics Theory · Mathematics 2019-10-25 Norbert Henze , M. Dolores Jiménez--Gamero

LLM-as-a-Judge evaluation has become a standard tool for assessing base model performance. However, characterizing performance via the naive estimator, i.e., raw judge outputs, is systematically biased. Recent work has proposed estimators…

Machine Learning · Computer Science 2026-05-11 James Fiedler

We present a new test of hypothesis in which we seek the probability of the null conditioned on the data, where the null is a simplification undertaken to counter the intractability of the more complex model, that the simpler null model is…

Applications · Statistics 2016-09-20 Dalia Chakrabarty

Robust classification algorithms have been developed in recent years with great success. We take advantage of this development and recast the classical two-sample test problem in the framework of classification. Based on the estimates of…

Statistics Theory · Mathematics 2019-09-18 Haiyan Cai , Bryan Goggin , Qingtang Jiang

The Jeffreys-Lindley paradox displays how the use of a p-value (or number of standard deviations z) in a frequentist hypothesis test can lead to an inference that is radically different from that of a Bayesian hypothesis test in the form…

Methodology · Statistics 2017-02-14 Robert D. Cousins

Estimation of the proportion of null hypotheses in a multiple testing problem can greatly enhance the performance of the existing algorithms. Although various estimators for the proportion of null hypotheses have been proposed, most are…

Statistics Theory · Mathematics 2024-09-09 Nabaneet Das