English
Related papers

Related papers: COWs and their Hybrids: A Statistical View of Cust…

200 papers

A common problem in data analysis is the separation of signal and background. We revisit and generalise the so-called $sWeights$ method, which allows one to calculate an empirical estimate of the signal density of a control variable using a…

Methodology · Statistics 2022-08-24 Hans Dembinski , Matthew Kenzie , Christoph Langenbruch , Michael Schmelling

In the multiple linear regression setting, we propose a general framework, termed weighted orthogonal components regression (WOCR), which encompasses many known methods as special cases, including ridge regression and principal components…

Machine Learning · Statistics 2018-01-24 Xiaogang Su , Yaa Wonkye , Pei Wang , Xiangrong Yin

A method is described, which computes from an observed sample of events upper limits for production rates of particles, or, in case of appearance of a signal, the probability for an upwards fluctuation of the background. For any candidate,…

High Energy Physics - Experiment · Physics 2010-10-27 P. Bock

Many scientific questions require estimating the effects of continuous treatments. Outcome modeling and weighted regression based on the generalized propensity score are the most commonly used methods to evaluate continuous effects.…

Methodology · Statistics 2019-10-29 Nathan Kallus , Michele Santacatterina

The challenge of Out-of-Distribution (OOD) generalization poses a foundational concern for the application of machine learning algorithms to risk-sensitive areas. Inspired by traditional importance weighting and propensity weighting…

Machine Learning · Computer Science 2025-02-12 Han Yu , Yue He , Renzhe Xu , Dongbai Li , Jiayin Zhang , Wenchao Zou , Peng Cui

Given a large set $U$ where each item $a\in U$ has weight $w(a)$, we want to estimate the total weight $W=\sum_{a\in U} w(a)$ to within factor of $1\pm\varepsilon$ with some constant probability $>1/2$. Since $n=|U|$ is large, we want to do…

Data Structures and Algorithms · Computer Science 2021-10-29 Lorenzo Beretta , Jakub Tětek

Composites, or linear combinations of variables, play an important role in multivariate behavioral research. They appear in the form of indices, inventories, formative constructs, parcels, and emergent variables. Although structural…

Methodology · Statistics 2025-09-03 Jörg Henseler , Xi Yu , Tamara Schamberger , Gregory R. Hancock , Florian Schuberth

Applied to statistical physics models, the random cost algorithm enforces a Random Walk (RW) in energy (or possibly other thermodynamic quantities). The dynamics of this procedure is distinct from fixed weight updates. The probability for a…

Statistical Mechanics · Physics 2009-10-31 Bernd A. Berg , Ulrich H. E. Hansmann

A common situation in experimental physics is to have a signal which can not be separated from a non-interfering background through the use of any cut. In this paper, we describe a procedure for determining, on an event-by-event basis, a…

Data Analysis, Statistics and Probability · Physics 2008-05-20 M. Williams , M. Bellis , C. A. Meyer

Forecast combination and model averaging have become popular tools in forecasting and prediction, both of which combine a set of candidate estimates with certain weights and are often shown to outperform single estimates. A data-driven…

Statistics Theory · Mathematics 2025-10-31 Jiahui Zou , Andrey Vasnev , Wendun Wang , Xinyu Zhang

Comparing structured data from possibly different metric-measure spaces is a fundamental task in machine learning, with applications in, e.g., graph classification. The Gromov-Wasserstein (GW) discrepancy formulates a coupling between the…

Machine Learning · Computer Science 2022-07-12 Hongwei Jin , Zishun Yu , Xinhua Zhang

In this paper, a new classifier based on the intrinsic properties of the data is proposed. Classification is an essential task in data mining-based applications. The classification problem will be challenging when the size of the training…

Machine Learning · Computer Science 2020-10-14 Sahar Tavakoli

Bias in causal comparisons has a direct correspondence with distributional imbalance of covariates between treatment groups. Weighting strategies such as inverse propensity score weighting attempt to mitigate bias by either modeling the…

Methodology · Statistics 2022-03-14 Jared D. Huling , Simon Mak

Sampling is a popular method for approximate inference when exact inference is impractical. Generally, sampling algorithms do not exploit context-specific independence (CSI) properties of probability distributions. We introduce…

Artificial Intelligence · Computer Science 2021-03-02 Nitesh Kumar , Ondřej Kuželka

Accurately estimating the proportion of true signals among a large number of variables is crucial for enhancing the precision and reliability of scientific research. Traditional signal proportion estimators often assume independence among…

Statistics Theory · Mathematics 2026-05-15 Jingtian Bai , Xinge Jessie Jeng

Traditional statistical mechanics is constrained by the binary paradigms of identical/distinguishable and bosonic/fermionic particle statistics, leading to a fundamental logical gap in describing systems with partial distinguishability. We…

Statistical Mechanics · Physics 2026-01-21 Wang Hao , Meng Yancen , Zhang Kuang , Zhou Rui'en

Traditional methods for covariate adjustment of treatment means in designed experiments are inherently conditional on the observed covariate values. In order to develop a coherent general methodology for analysis of covariance, we propose a…

Methodology · Statistics 2010-01-19 James G. Booth , Walter T. Federer , Martin T. Wells , Russell D. Wolfinger

Covariate balance is crucial in obtaining unbiased estimates of treatment effects in observational studies. Methods based on inverse probability weights have been widely used to estimate treatment effects with observational data. Machine…

Methodology · Statistics 2021-04-08 Michele Santacatterina

Marginal structural models (MSMs) estimate the causal effect of a time-varying treatment in the presence of time-dependent confounding via weighted regression. The standard approach of using inverse probability of treatment weighting (IPTW)…

Methodology · Statistics 2019-08-13 Nathan Kallus , Michele Santacatterina

Global expression analyses using microarray technologies are becoming more common in genomic research, therefore, new statistical challenges associated with combining information from multiple studies must be addressed. In this paper we…

Applications · Statistics 2013-01-29 Jia Li , George C. Tseng
‹ Prev 1 2 3 10 Next ›