中文
相关论文

相关论文: IQ: Intrinsic measure for quantifying the heteroge…

200 篇论文

A measure of dependence is said to be equitable if it gives similar scores to equally noisy relationships of different types. Equitability is important in data exploration when the goal is to identify a relatively small set of strongest…

机器学习 · 计算机科学 2013-08-16 David Reshef , Yakir Reshef , Michael Mitzenmacher , Pardis Sabeti

Imbalance in classification tasks is commonly quantified by the cardinalities of examples across classes. This, however, disregards the presence of redundant examples and inherent differences in the learning difficulties of classes.…

机器学习 · 计算机科学 2026-01-22 Çağrı Eser , Zeynep Sonat Baltacı , Emre Akbaş , Sinan Kalkan

Measuring interdisciplinarity is a pertinent but challenging issue in quantitative studies of science. There seems to be a consensus in the literature that the concept of interdisciplinarity is multifaceted and ambiguous. Unsurprisingly,…

数字图书馆 · 计算机科学 2019-12-11 Qi Wang , Jesper Wiborg Schneider

Meta-analysis is a statistical method to combine results from multiple clinical or genomic studies with the same or similar research problems. It has been widely use to increase statistical power in finding clinical or genomic differences…

统计理论 · 数学 2019-08-05 Yusi Fang , Shaowu Tang , Zhiguang Huo , George C. Tseng , Yongseok Park

This paper is concerned with the study of constrained statistical learning problems, the unconstrained version of which are at the core of virtually all of modern information processing. Accounting for constraints, however, is paramount to…

机器学习 · 计算机科学 2020-02-14 Luiz F. O. Chamon , Santiago Paternain , Miguel Calvo-Fullana , Alejandro Ribeiro

There is a growing need for flexible general frameworks that integrate individual-level data with external summary information for improved statistical inference. External information relevant for a risk prediction model may come in…

统计方法学 · 统计学 2023-04-11 Tian Gu , Jeremy M. G. Taylor , Bhramar Mukherjee

Assessment of multimedia quality relies heavily on subjective assessment, and is typically done by human subjects in the form of preferences or continuous ratings. Such data is crucial for analysis of different multimedia processing…

多媒体 · 计算机科学 2018-01-26 Manish Narwaria , Lukas Krasula , Patrick Le Callet

Data depth has been applied as a nonparametric measurement for ranking multivariate samples. In this paper, we focus on homogeneity tests to assess whether two multivariate samples are from the same distribution. There are many data…

统计理论 · 数学 2023-06-09 Yiting Chen , Wei Lin , Xiaoping Shi

Empirical claims often rely on one population, design, and analysis. Many-analysts, multiverse, and robustness studies expose how results can vary across plausible analytic choices. Synthesizing these results, however, is nontrivial as all…

统计方法学 · 统计学 2025-11-24 František Bartoš , Suzanne Hoogeveen , Alexandra Sarafoglou , Samuel Pawel

Comparative binary outcome data are of fundamental interest in statistics and are often pooled in meta-analyses. Here we examine the simplest case where for each study there are two patient groups and a binary event of interest, giving rise…

统计方法学 · 统计学 2018-06-12 Rose Baker , Dan Jackson

Information-theoretic (IT) measures are ubiquitous in artificial intelligence: entropy drives decision-tree splits and uncertainty quantification, cross-entropy is the default classification loss, mutual information underpins representation…

人工智能 · 计算机科学 2026-04-28 Nikolaos Al. Papadopoulos , Konstantinos E. Psannis

Cochran's $Q$ statistic is routinely used for testing heterogeneity in meta-analysis. Its expected value (under an incorrect null distribution) is part of several popular estimators of the between-study variance, $\tau^2$. Those…

统计方法学 · 统计学 2023-04-11 Elena Kulinskaya , David C. Hoaglin

In this paper, we propose standard statistical tools as a solution to commonly highlighted problems in the explainability literature. Indeed, leveraging statistical estimators allows for a proper definition of explanations, enabling…

机器学习 · 统计学 2024-05-01 Valentina Ghidini

Heteroskedasticity is a statistical anomaly that describes differing variances of error terms in a time series dataset. The presence of heteroskedasticity in data imposes serious challenges for forecasting models and many statistical tests…

统计理论 · 数学 2016-09-21 Marwa Hassan , Mo Hossny , Douglas Creighton , Saeid Nahavandi

Recent years have seen the development of many novel scoring tools for disease prognosis and prediction. To become accepted for use in clinical applications, these tools have to be validated on external data. In practice, validation is…

统计方法学 · 统计学 2022-12-06 Matthias Schmid , Tim Friede , Nadja Klein , Leonie Weinhold

We propose a test-based elastic integrative analysis of the randomized trial and real-world data to estimate treatment effect heterogeneity with a vector of known effect modifiers. When the real-world data are not subject to bias, our…

统计方法学 · 统计学 2022-11-30 Shu Yang , Chenyin Gao , Donglin Zeng , Xiaofei Wang

It is increasingly common to collect pre-post data with pseudonyms or self-constructed identifiers. On survey responses from sensitive populations, identifiers may be made optional to encourage higher response rates. The ability to match…

统计方法学 · 统计学 2023-12-25 Raymond Pomponio , Bailey K. Fosdick , Julia Wrobel , Ryan A. Peterson

Self-consistency-based approaches, which involve repeatedly sampling multiple outputs and selecting the most consistent one as the final response, prove to be remarkably effective in improving the factual accuracy of large language models.…

Many practical studies rely on hypothesis testing procedures applied to data sets with missing information. An important part of the analysis is to determine the impact of the missing data on the performance of the test, and this can be…

统计方法学 · 统计学 2011-02-15 Dan L. Nicolae , Xiao-Li Meng , Augustine Kong

The inference of causal relationships using observational data from partially observed multivariate systems with hidden variables is a fundamental question in many scientific domains. Methods extracting causal information from conditional…

机器学习 · 统计学 2020-10-13 Daniel Chicharro , Michel Besserve , Stefano Panzeri