English
Related papers

Related papers: Bounds for the weight of external data in shrinkag…

200 papers

Machine learning models have traditionally been developed under the assumption that the training and test distributions match exactly. However, recent success in few-shot learning and related problems are encouraging signs that these models…

Machine Learning · Statistics 2020-10-15 James Lucas , Mengye Ren , Irene Kameni , Toniann Pitassi , Richard Zemel

Recent years have seen the development of many novel scoring tools for disease prognosis and prediction. To become accepted for use in clinical applications, these tools have to be validated on external data. In practice, validation is…

Methodology · Statistics 2022-12-06 Matthias Schmid , Tim Friede , Nadja Klein , Leonie Weinhold

We develop a method for hybrid analyses that uses external controls to augment internal control arms in randomized controlled trials (RCT) where the degree of borrowing is determined based on similarity between RCT and external control…

Methodology · Statistics 2023-05-11 Evan Kwiatkowski , Jiawen Zhu , Xiao Li , Herbert Pang , Grazyna Lieberman , Matthew A. Psioda

Longitudinal data are common in clinical trials and observational studies, where missing outcomes due to dropouts are always encountered. Under such context with the assumption of missing at random, the weighted generalized estimating…

Methodology · Statistics 2019-04-30 Chixiang Chen , Biyi Shen , Lijun Zhang , Yuan Xue , Ming Wang

Several statistical and machine learning methods are proposed to estimate the type and intensity of physical load and accumulated fatigue . They are based on the statistical analysis of accumulated and moving window data subsets with…

Machine Learning · Computer Science 2018-12-11 Sergii Stirenko , Gang Peng , Wei Zeng , Yuri Gordienko , Oleg Alienin , Oleksandr Rokovyi , Nikita Gordienko

State-of-the-art methods for conditional average treatment effect (CATE) estimation make widespread use of representation learning. Here, the idea is to reduce the variance of the low-sample CATE estimation by a (potentially constrained)…

Machine Learning · Statistics 2026-03-13 Valentyn Melnychuk , Dennis Frauen , Stefan Feuerriegel

Heterogeneous treatment effects (HTEs) are increasingly estimated using machine learning models that produce highly personalized predictions of treatment effects. In practice, however, predicted treatment effects are rarely interpreted,…

Researchers frequently estimate treatment effects by regressing outcomes (Y) on treatment (D) and covariates (X). Even without unobserved confounding, the coefficient on D yields a conditional-variance-weighted average of strata-wise…

Methodology · Statistics 2025-05-05 Tanvi Shinkre , Chad Hazlett

A new meta-algorithm for estimating the conditional average treatment effects is proposed in the paper. The main idea underlying the algorithm is to consider a new dataset consisting of feature vectors produced by means of concatenation of…

Machine Learning · Statistics 2019-09-10 Lev V. Utkin , Mikhail V. Kots , Viacheslav S. Chukanov

The prediction interval has been increasingly used in meta-analyses as a useful measure for assessing the magnitude of treatment effect and between-studies heterogeneity. In calculations of the prediction interval, although the…

Methodology · Statistics 2021-07-14 Yuta Hamaguchi , Hisashi Noma , Kengo Nagashima , Tomohide Yamada , Toshi A. Furukawa

Multivariate meta-analysis can be adapted to a wide range of situations for multiple outcomes and multiple treatment groups when combining studies together. The within-study correlation between effect sizes is often assumed known in…

Methodology · Statistics 2020-03-12 Xiaohuan Xue

Meta-analysis is a powerful tool to synthesize findings from multiple studies. The normal-normal random-effects model is widely used to account for between-study heterogeneity. However, meta-analysis of sparse data, which may arise when the…

Methodology · Statistics 2024-06-10 Taojun Hu , Yi Zhou , Satoshi Hattori

Modern statistics provides an ever-expanding toolkit for estimating unknown parameters. Consequently, applied statisticians frequently face a difficult decision: retain a parameter estimate from a familiar method or replace it with an…

Methodology · Statistics 2022-12-20 Brian L. Trippe , Sameer K. Deshpande , Tamara Broderick

We tackle covariance estimation in low-sample scenarios, employing a structured covariance matrix with shrinkage methods. These involve convexly combining a low-bias/high-variance empirical estimate with a biased regularization estimator,…

Instrumentation and Methods for Astrophysics · Physics 2024-06-28 Olivier Flasseur , Eric Thiébaut , Loïc Denis , Maud Langlois

Modern data is messy and high-dimensional, and it is often not clear a priori what are the right questions to ask. Instead, the analyst typically needs to use the data to search for interesting analyses to perform and hypotheses to test.…

Machine Learning · Statistics 2019-10-09 Daniel Russo , James Zou

We study the basic task of mean estimation in the presence of mean-shift contamination. In the mean-shift contamination model, an adversary is allowed to replace a small constant fraction of the clean samples by samples drawn from…

Machine Learning · Computer Science 2026-02-27 Ilias Diakonikolas , Giannis Iakovidis , Daniel M. Kane , Sihan Liu

Weighting methods are used in observational studies to adjust for covariate imbalances between treatment and control groups. Entropy balancing (EB) is an alternative to inverse probability weighting with an estimated propensity score. The…

Methodology · Statistics 2022-04-25 David Källberg , Ingeborg Waernbaum

The following problem is considered: given a joint distribution $P_{XY}$ and an event $E$, bound $P_{XY}(E)$ in terms of $P_XP_Y(E)$ (where $P_XP_Y$ is the product of the marginals of $P_{XY}$) and a measure of dependence of $X$ and $Y$.…

Information Theory · Computer Science 2019-03-12 Ibrahim Issa , Amedeo Roberto Esposito , Michael Gastpar

As learning difficulty is crucial for machine learning (e.g., difficulty-based weighting learning strategies), previous literature has proposed a number of learning difficulty measures. However, no comprehensive investigation for learning…

Machine Learning · Computer Science 2022-09-20 Weiyao Zhu , Ou Wu , Fengguang Su , Yingjun Deng

In various biomedical studies, analysis often focuses on data magnitudes, particularly when algebraic signs are irrelevant or lost. For repeated measures studies involving magnitude outcomes, incorporating random effects is essential as…

Methodology · Statistics 2025-07-16 Wen Teng , Niall D. Ferguson , Ewan C. Goligher , Anna Heath
‹ Prev 1 8 9 10 Next ›