中文
相关论文

相关论文: Scale two-sample testing with arbitrarily missing …

200 篇论文

Missing values are unavoidable in many applications of machine learning and present challenges both during training and at test time. When variables are missing in recurring patterns, fitting separate pattern submodels have been proposed as…

机器学习 · 计算机科学 2023-11-27 Lena Stempfle , Ashkan Panahi , Fredrik D. Johansson

Combining dependent p-values poses a long-standing challenge in statistical inference, particularly when aggregating findings from multiple methods to enhance signal detection. Recently, p-value combination tests based on regularly…

统计方法学 · 统计学 2025-04-22 Lin Gui , Yuchao Jiang , Jingshu Wang

The focus of this study is to evaluate the effectiveness of Machine Learning (ML) methods for two-sample testing with right-censored observations. To achieve this, we develop several ML-based methods with varying architectures and implement…

机器学习 · 计算机科学 2024-09-27 Petr Philonenko , Sergey Postovalov

This paper focuses on a data-rich environment where the data set has a very large cross-sectional dimension, is likely to exhibit local dependence, and yet is hard to determine the dependence ordering. Such a situation arises, for example,…

统计方法学 · 统计学 2018-07-03 Kyungchul Song

We propose a new and computationally efficient algorithm for maximizing the observed log-likelihood for a multivariate normal data matrix with missing values. We show that our procedure based on iteratively regressing the missing on the…

统计方法学 · 统计学 2012-11-21 Nicolas Städler , Daniel J. Stekhoven , Peter Bühlmann

This paper addresses detecting anomalous patterns in images, time-series, and tensor data when the location and scale of the pattern is unknown a priori. The multiscale scan statistic convolves the proposed pattern with the image at various…

统计理论 · 数学 2018-06-22 James Sharpnack

This article is concerned with simultaneous tests on linear regression coefficients in high-dimensional settings. When the dimensionality is larger than the sample size, the classic $F$-test is not applicable since the sample covariance…

统计方法学 · 统计学 2015-02-17 Long Feng

Testing the equality of the covariance matrices of two high-dimensional samples is a fundamental inference problem in statistics. Several tests have been proposed but they are either too liberal or too conservative when the required…

统计理论 · 数学 2023-01-04 Jin-Ting Zhang , Jingyi Wang , Tianming Zhu

Large-scale multiple two-sample {\em Student}'s $t$ testing problems often arise from the statistical analysis of scientific data. To detect components with different values between two mean vectors, a well-known procedure is to apply the…

统计方法学 · 统计学 2014-10-17 Weidong Liu

Logistic regression is widely used to model the propensity score in the analysis of nonignorable missing data. However, goodness-of-fit testing for this propensity score model has received limited attention in the literature. In this paper,…

统计方法学 · 统计学 2026-04-24 Manli Cheng , Yangjianchen Xu , Qinglong Tian , Pengfei Li

In this paper, we introduce a new method for testing the stationarity of time series, where the test statistic is obtained from measuring and maximising the difference in the second-order structure over pairs of randomly drawn intervals.…

统计方法学 · 统计学 2016-11-29 Haeran Cho

The goal of two-sample tests is to assess whether two samples, $S_P \sim P^n$ and $S_Q \sim Q^m$, are drawn from the same distribution. Perhaps intriguingly, one relatively unexplored method to build two-sample tests is the use of binary…

机器学习 · 统计学 2018-03-14 David Lopez-Paz , Maxime Oquab

In medical domain, data features often contain missing values. This can create serious bias in the predictive modeling. Typical standard data mining methods often produce poor performance measures. In this paper, we propose a new method to…

机器学习 · 统计学 2015-03-24 Talayeh Razzaghi , Oleg Roderick , Ilya Safro , Nick Marko

Predictive values are measures of the clinical accuracy of a binary diagnostic test, and depend on the sensitivity and the specificity of the test and on the disease prevalence among the population being studied. This article studies…

其他统计学 · 统计学 2024-08-14 Jose Antonio Roldan-Nofuentes

Clustered competing risks data are commonly encountered in multicenter studies. The analysis of such data is often complicated due to informative cluster size, a situation where the outcomes under study are associated with the size of the…

统计方法学 · 统计学 2021-04-26 Wenxian Zhou , Giorgos Bakoyannis , Ying Zhang , Constantin T. Yiannoutsos

In a high dimensional regression setting in which the number of variables ($p$) is much larger than the sample size ($n$), the number of possible two-way interactions between the variables is immense. If the number of variables is in the…

统计方法学 · 统计学 2024-06-26 Marianne A Jonker , Luc van Schijndel , Eric Cator

Hypothesis testing is a statistical inference approach used to determine whether data supports a specific hypothesis. An important type is the two-sample test, which evaluates whether two sets of data points are from identical…

机器学习 · 计算机科学 2025-01-08 Weizhi Li , Visar Berisha , Gautam Dasarathy

Missing data occur frequently in a wide range of applications. In this paper, we consider estimation of high-dimensional covariance matrices in the presence of missing observations under a general missing completely at random model in the…

统计方法学 · 统计学 2016-05-17 T. Tony Cai , Anru Zhang

We present a robust test for change-points in time series which is based on the two-sample Hodges-Lehmann estimator. We develop new limit theory for a class of statistics based on the two-sample U-quantile processes, in the case of short…

统计理论 · 数学 2019-05-17 Herold Dehling , Roland Fried , Martin Wendler

We propose new statistical tests, in high-dimensional settings, for testing the independence of two random vectors and their conditional independence given a third random vector. The key idea is simple, i.e., we first transform each…

统计方法学 · 统计学 2026-01-28 Jinyuan Chang , Yue Du , Jing He , Qiwei Yao
‹ 上一页 1 8 9 10 下一页 ›