中文
相关论文

相关论文: On the Impossibility of Specification Testing of I…

200 篇论文

The paper presents some models for the propensity score. Considerable attention is given to a recently popular, but relatively under-explored setting in causal inference where the no-interference assumption does not hold. We lay out some…

统计方法学 · 统计学 2022-08-16 Hyunseung Kang , Chan Park , Ralph Trane

Although overparameterized models have shown their success on many machine learning tasks, the accuracy could drop on the testing distribution that is different from the training one. This accuracy drop still limits applying machine…

机器学习 · 计算机科学 2022-09-29 Yiping Lu , Wenlong Ji , Zachary Izzo , Lexing Ying

Compositionality supports the manipulation of large systems by working on their components. For model-based testing, this means that large systems can be tested by modelling and testing their components: passing tests for all components…

软件工程 · 计算机科学 2025-08-01 Gijs van Cuyck , Lars van Arragon , Jan Tretmans

If an experimental treatment is experienced by both treated and control group units, tests of hypotheses about causal effects may be difficult to conceptualize let alone execute. In this paper, we show how counterfactual causal models may…

统计方法学 · 统计学 2012-08-03 Jake Bowers , Mark Fredrickson , Costas Panagopoulos

Benchmarking studies in computational chemistry use reference datasets to assess the accuracy of a method through error statistics. The commonly used error statistics, such as the mean signed and mean unsigned errors, do not inform…

化学物理 · 物理学 2018-03-19 Pascal Pernot , Andreas Savin

Large-scale simultaneous hypothesis testing appears in many areas such as microarray studies, genome-wide association studies, brain imaging, disease mapping and astronomical surveys. A well-known inference method is to control the false…

统计方法学 · 统计学 2025-07-22 Xiaoqing Niu , Pengfei Li , Yuejiao Fu

Environmental epidemiologists are increasingly interested in establishing causality between exposures and health outcomes. A popular model for causal inference is the Rubin Causal Model (RCM), which typically seeks to estimate the average…

应用统计 · 统计学 2021-01-26 Keith W. Zirkle , Marie-Abele Bind , Jenise L. Swall , David C. Wheeler

Public health researchers often estimate health effects of exposures (e.g., pollution, diet, lifestyle) that cannot be directly measured for study subjects. A common strategy in environmental epidemiology is to use a first-stage (exposure)…

统计方法学 · 统计学 2014-06-03 Adam A. Szpiro , Christopher J. Paciorek

Score-based diffusion models are a powerful class of generative models, but their practical use often depends on training neural networks to approximate the score function. Training-free diffusion models provide an attractive alternative by…

数值分析 · 数学 2026-01-28 Pengjun Wang , Zezhong Zhang , Minglei Yang , Feng Bao , Yanzhao Cao , Guannan Zhang

It is generally believed that any particle to be discovered will have a TeV-order mass. Given its great mass, it must have a large decay width. Therefore, the interference effect will be very common if they and the Standard-Model (SM)…

高能物理 - 唯象学 · 物理学 2019-11-27 Li-Gang Xia

Software Product Lines (SPL) are inherently difficult to test due to the combinatorial explosion of the number of products to consider. To reduce the number of products to test, sampling techniques such as combinatorial interaction testing…

软件工程 · 计算机科学 2017-10-24 Xavier Devroey , Maxime Cordy , Gilles Perrouin , Pierre-Yves Schobbens , Axel Legay , Patrick Heymans

This paper shows that the problem of testing hypotheses in moment condition models without any assumptions about identification may be considered as a problem of testing with an infinite-dimensional nuisance parameter. We introduce a…

统计理论 · 数学 2014-09-24 Isaiah Andrews , Anna Mikusheva

Experimental research on behavior and cognition frequently rests on stimulus or subject selection where not all characteristics can be fully controlled, even when attempting strict matching. For example, when contrasting patients to…

统计方法学 · 统计学 2016-08-29 Jona Sassenhagen , Phillip M. Alday

Traditional hypothesis tests for differences between binomial proportions are at risk of being too liberal (Wald test) or overly conservative (Fisher's exact test). This problem is exacerbated in small samples. Regulators favour exact…

统计方法学 · 统计学 2025-07-31 Stef Baas , Yaron Racah , Elad Berkman , Sofia S. Villar

Motivated by real-world machine learning applications, we analyze approximations to the non-asymptotic fundamental limits of statistical classification. In the binary version of this problem, given two training sequences generated according…

信息论 · 计算机科学 2018-12-07 Lin Zhou , Vincent Y. F. Tan , Mehul Motani

In the context of finite mixture models one considers the problem of classifying as many observations as possible in the classes of interest while controlling the classification error rate in these same classes. Similar to what is done in…

机器学习 · 计算机科学 2021-09-30 Tristan Mary-Huard , Vittorio Perduca , Gilles Blanchard , Martin-Magniette Marie-Laure

In model checking for regressions, nonparametric estimation-based tests usually have tractable limiting null distributions and are sensitive to oscillating alternative models, but suffer from the curse of dimensionality. In contrast,…

统计方法学 · 统计学 2019-03-12 Lingzhu Li , Xuehu Zhu , Lixing Zhu

The repeated community-wide reuse of test sets in popular benchmark problems raises doubts about the credibility of reported test-error rates. Verifying whether a learned model is overfitted to a test set is challenging as independent test…

机器学习 · 计算机科学 2019-11-15 Roman Werpachowski , András György , Csaba Szepesvári

Measurement error arises through a variety of mechanisms. A rich literature exists on the bias introduced by covariate measurement error and on methods of analysis to address this bias. By comparison, less attention has been given to errors…

统计方法学 · 统计学 2018-11-27 Pamela Shaw , Jiwei He , Bryan Shepherd

Statistical significance testing of differences in values of metrics like recall, precision and balanced F-score is a necessary part of empirical natural language processing. Unfortunately, we find in a set of experiments that many commonly…

计算与语言 · 计算机科学 2007-05-23 Alexander Yeh