English
Related papers

Related papers: Differential Test Functioning via Robust Scaling

200 papers

Item Response Theory becomes an increasingly important tool when analyzing ``Big Data'' gathered from online educational venues. However, the mechanism was originally developed in traditional exam settings, and several of its assumptions…

Physics Education · Physics 2014-05-29 Gerd Kortemeyer

Motivated by recent data analytics applications, we study the adversarial robustness of robust estimators. Instead of assuming that only a fraction of the data points are outliers as considered in the classic robust estimation setup, in…

Statistics Theory · Mathematics 2020-04-01 Lifeng Lai , Erhan Bayraktar

In a group testing scheme, a set of tests is designed to identify a small number $t$ of defective items that are present among a large number $N$ of items. Each test takes as input a group of items and produces a binary output indicating…

Information Theory · Computer Science 2016-11-17 Arya Mazumdar

In many Deep Reinforcement Learning (RL) problems, decisions in a trained policy vary in significance for the expected safety and performance of the policy. Since RL policies are very complex, testing efforts should concentrate on states in…

Machine Learning · Computer Science 2024-11-13 Stefan Pranger , Hana Chockler , Martin Tappler , Bettina Könighofer

Adversarial examples, generated by carefully crafted perturbation, have attracted considerable attention in research fields. Recent works have argued that the existence of the robust and non-robust features is a primary cause of the…

Machine Learning · Computer Science 2022-04-07 Junho Kim , Byung-Kwan Lee , Yong Man Ro

There is growing interest in developing causal inference methods for multi-valued treatments with a focus on pairwise average treatment effects. Here we focus on a clinically important, yet less-studied estimand: causal drug-drug…

Methodology · Statistics 2022-09-20 Di Shu , Peisong Han , Sean Hennessy , Todd A Miano

Most work on one-shot devices assume that there is only one possible cause of device failure. However, in practice, it is often the case that the products under study can experience any one of various possible causes of failure. Robust…

Applications · Statistics 2020-04-29 N. Balakrishnan , E. Castilla , N. Martin , L. Pardo

Test-time defenses are used to improve the robustness of deep neural networks to adversarial examples during inference. However, existing methods either require an additional trained classifier to detect and correct the adversarial samples,…

Machine Learning · Computer Science 2024-08-26 Anurag Singh , Mahalakshmi Sabanayagam , Krikamol Muandet , Debarghya Ghoshdastidar

In certain academic systems, a student can enroll for an exam immediately after the end of the teaching period or can postpone it to any later examination session, so that the grade is missing until the exam is not attempted. We propose an…

Methodology · Statistics 2016-09-22 Silvia Bacci , Francesco Bartolucci , Leonardo Grilli , Carla Rampichini

We propose a framework to construct practical kernel-based two-sample tests from the family of $f$-divergences. The test statistic is computed from the witness function of a regularized variational representation of the divergence, which we…

Machine Learning · Statistics 2026-01-28 Mónica Ribero , Antonin Schrab , Arthur Gretton

We propose a function-valued evaluation metric for generative models based on the relative density ratio (RDR) designed to characterize distributional differences between real and generated samples. As an evaluation metric, the RDR function…

Methodology · Statistics 2025-12-29 Yuliang Xu , Yun Wei , Li Ma

The difference-in-differences (DID) method identifies the average treatment effects on the treated (ATT) under mainly the so-called parallel trends (PT) assumption. The most common and widely used approach to justify the PT assumption is…

Econometrics · Economics 2023-08-23 Kyunghoon Ban , Désiré Kédagni

At present there are two vastly different ab initio approaches to the description of the the many-body dynamics: the Density Functional Theory (DFT) and the functional integral (path integral) approaches. On one hand, if implemented…

Nuclear Theory · Physics 2014-11-20 Aurel Bulgac

The disparity in accuracy between classes in standard training is amplified during adversarial training, a phenomenon termed the robust fairness problem. Existing methodologies aimed to enhance robust fairness by sacrificing the model's…

Machine Learning · Computer Science 2024-01-24 Hyungyu Lee , Saehyung Lee , Hyemi Jang , Junsung Park , Ho Bae , Sungroh Yoon

Errors might not have the same consequences depending on the task at hand. Nevertheless, there is limited research investigating the impact of imbalance in the contribution of different features in an error vector. Therefore, we propose the…

Machine Learning · Computer Science 2022-07-12 Xavier F. Cadet , Sara Ahmadi-Abhari , Hamed Haddadi

There is a well-known problem in Null Hypothesis Significance Testing: many statistically significant results fail to replicate in subsequent experiments. We show that this problem arises because standard `point-form null' significance…

Methodology · Statistics 2025-02-06 Fintan Costello , Paul Watts

Nuclear density functional theory (DFT) is the only microscopic, global approach to the structure of atomic nuclei. It is used in numerous applications, from determining the limits of stability to gaining a deep understanding of the…

Nuclear Theory · Physics 2015-02-06 Nicolas Schunck , Jordan D. McDonnell , Jason Sarich , Stefan M. Wild , Dave Higdon

We use variation of test scores measuring closely related skills to isolate peer effects. The intuition for our identification strategy is that the difference in closely related scores eliminates factors common to the performance in either…

General Economics · Economics 2025-07-03 Guido Kuersteiner , Ingmar Prucha , Ying Zeng

This paper considers identification and estimation of causal effect parameters from participating in a binary treatment in a difference in differences (DID) setup when the parallel trends assumption holds after conditioning on observed…

Econometrics · Economics 2024-06-25 Carolina Caetano , Brantly Callaway , Stroud Payne , Hugo Sant'Anna Rodrigues

Adaptive experiments, including efficient average treatment effect estimation and multi-armed bandit algorithms, have garnered attention in various applications, such as social experiments, clinical trials, and online advertisement…

Methodology · Statistics 2021-03-24 Masahiro Kato