English
Related papers

Related papers: Valid t-ratio Inference for IV

200 papers

Consider a one-way analysis of covariance model. Suppose that the parameter of interest theta is a specified linear contrast of the expected responses, for a given value of the covariate. Also suppose that the inference of interest is a…

Methodology · Statistics 2017-10-18 Waruni Abeysekera , Paul Kabaila , Oguzhan Yilmaz

The ratio of Bayesian evidences is a popular tool in cosmology to compare different models. There are however several issues with this method: Bayes' ratio depends on the prior even in the limit of non-informative priors, and Jeffrey's…

Cosmology and Nongalactic Astrophysics · Physics 2024-12-16 Luca Amendola , Vrund Patel , Ziad Sakr , Elena Sellentin , Kevin Wolz

We study conditions under which the addition of variables to a regression equation can turn a previously statistically insignificant result into a significant one. Specifically, we characterize the minimum strength of association required…

Statistics Theory · Mathematics 2025-09-24 Danielle Tsao , Ronan Perry , Carlos Cinelli

We consider the issue of performing accurate small-sample testing inference in beta regression models, which are useful for modeling continuous variates that assume values in $(0,1)$, such as rates and proportions. We derive the Bartlett…

Methodology · Statistics 2015-01-30 Fábio M. Bayer , Francisco Cribari-Neto

This paper introduces novel weighted conformal p-values and methods for model-free selective inference. The problem is as follows: given test units with covariates $X$ and missing responses $Y$, how do we select units for which the…

Methodology · Statistics 2023-09-27 Ying Jin , Emmanuel J. Candès

We study the properties of several likelihood-based statistics commonly used in testing for the presence of a known signal under a mixture model with known background, but unknown signal fraction. Under the null hypothesis of no signal, all…

Data Analysis, Statistics and Probability · Physics 2018-12-26 Igor Volobouev , A. Alexandre Trindade

When many (m) null hypotheses are tested with a single dataset, the control of the number of false rejections is often the principal consideration. Two popular controlling rates are the probability of making at least one false discovery…

Methodology · Statistics 2013-07-11 Djalel Eddine Meskaldji , Jean-Philippe Thiran , Stephan Morgenthaler

TF-IDF is a classical formula that is widely used for identifying important terms within documents. We show that TF-IDF-like scores arise naturally from the test statistic of a penalized likelihood-ratio test setup capturing word burstiness…

Computation and Language · Computer Science 2026-04-07 Zeyad Ahmed , Paul Sheridan , Michael McIsaac , Aitazaz A. Farooque

To enhance the reasoning capabilities of Large Language Models (LLMs) without high costs of training, nor extensive test-time sampling, we introduce Verification-First (VF), a strategy that prompts models to verify a provided candidate…

Computation and Language · Computer Science 2026-05-26 Shiguang Wu , Quanming Yao

The complex nature of inertial confinement fusion (ICF) experiments results in a very large number of experimental parameters that are only known with limited reliability. These parameters, combined with the myriad physical models that…

Plasma Physics · Physics 2015-06-15 Jim A Gaffney , Dan Clark , Vijay Sonnad , Stephen B Libby

Transitioning from Phase 2 to Phase 3 in drug development, at a rate of $\approx$40%, is the most stringent among phase transitions (Hay et al. (2014)). Yet, success rate at Phase 3 leading to approval is only $\approx$50% (Arrowsmith…

Methodology · Statistics 2025-10-29 Yujia Sun , Yang Han , Xingya Wang , Szu-Yu Tang , Yushi Liu , Jason C. Hsu

Practical or scientific considerations often lead to selecting a subset of parameters as ``important.'' Inferences about those parameters often are based on the same data used to select them in the first place. That can make the reported…

Methodology · Statistics 2019-06-04 Yoav Benjamini , Yotam Hechtlinger , Philip B. Stark

As large language models (LLMs) are increasingly deployed in critical decision-making systems, the lack of reliable methods to measure their uncertainty presents a fundamental trustworthiness risk. We introduce a normalized confidence score…

Machine Learning · Computer Science 2026-03-10 Xie Xiaohu , Liu Xiaohu , Yao Benjamin

Quantiles can represent key operational and business metrics, but the computational challenges associated with inference has hampered their adoption in online experimentation. One-sample confidence intervals are trivial to construct;…

Methodology · Statistics 2024-08-02 Evan Miller

Two recently introduced model based bias corrected estimators for proportion of true null hypotheses ($\pi_0$) under multiple hypotheses testing scenario have been restructured for exponentially distributed random observations available for…

Statistics Theory · Mathematics 2020-07-28 Aniket Biswas , Gaurangadeb Chattopadhyay , Aditya Chatterjee

In a context of multiple hypothesis testing, we provide several new exact calculations related to the false discovery proportion (FDP) of step-up and step-down procedures. For step-up procedures, we show that the number of erroneous…

Statistics Theory · Mathematics 2011-06-29 Etienne Roquain , Fanny Villers

Feature Selection (FS) under domain adaptation (DA) is a critical task in machine learning, especially when dealing with limited target data. However, existing methods lack the capability to guarantee the reliability of FS under DA. In this…

Machine Learning · Statistics 2024-10-22 Nguyen Thang Loi , Duong Tan Loc , Vo Nguyen Le Duy

Adjustment of statistical significance levels for repeated analysis in group sequential trials has been understood for some time. Similarly, methods for adjustment accounting for testing multiple hypotheses are common. There is limited…

Methodology · Statistics 2023-11-28 Yujie Zhao , Qi Liu , Linda Z. Sun , Keaven M. Anderson

Gaussian conditional realizations are routinely used for risk assessment and planning in a variety of Earth sciences applications. Conditional realizations can be obtained by first creating unconditional realizations that are then…

Methodology · Statistics 2017-02-07 Denis Marcotte , Denis Allard

Hybrid clinical trials, that borrow real-world data (RWD), are gaining interest, especially for rare diseases. They assume RWD and randomized control arm be exchangeable, but violations can bias results, inflate type I error, or reduce…