中文
相关论文

相关论文: Computing AIC for black-box models using Generalis…

200 篇论文

Our goal is to evaluate the accuracy of a black-box classification model, not as a single aggregate on a given test data distribution, but as a surface over a large number of combinations of attributes characterizing multiple test data…

机器学习 · 计算机科学 2021-10-27 Vihari Piratla , Soumen Chakrabarty , Sunita Sarawagi

The integration of Artificial Intelligence (AI) in Network Intrusion Detection Systems (NIDS) is a promising approach to tackle the increasing sophistication of cyberattacks. However, since Machine Learning (ML) and Deep Learning (DL)…

密码学与安全 · 计算机科学 2025-11-13 Miguel Silva , Daniela Pinto , João Vitorino , Eva Maia , Isabel Praça , Ivone Amorim , Maria João Viamonte

Two key tasks in high-dimensional regularized regression are tuning the regularization strength for accurate predictions and estimating the out-of-sample risk. It is known that the standard approach -- $k$-fold cross-validation -- is…

统计理论 · 数学 2025-10-24 Kevin Luo , Yufan Li , Pragya Sur

In this paper, conditional data augmentation (DA) is investigated for the degrees of freedom parameter $\nu$ of a Student-$t$ distribution. Based on a restricted version of the expected augmented Fisher information, it is conjectured that…

统计方法学 · 统计学 2021-09-07 Darjus Hosszejni

For prediction models developed on clustered data that do not account for cluster heterogeneity in model parameterization, it is crucial to use cluster-based validation to assess model generalizability on unseen clusters. This paper…

统计方法学 · 统计学 2025-06-23 Jiaxing Qiu , Douglas E. Lake , Pavel Chernyavskiy , Teague R. Henry

Classification models play a central role in data-driven decision-making applications such as medical diagnosis, recommendation systems, and risk assessment. Traditional performance metrics, such as accuracy and AUC, focus on overall error…

机器学习 · 计算机科学 2026-04-03 Chen Yang , Zheng Cui , Daniel Zhuoyu Long , Jin Qi , Ruohan Zhan

Causality has been combined with machine learning to produce robust representations for domain generalization. Most existing methods of this type require massive data from multiple domains to identify causal features by cross-domain…

机器学习 · 计算机科学 2024-03-01 Yang Chen , Yitao Liang , Zhouchen Lin

Most fair machine learning methods either highly rely on the sensitive information of the training samples or require a large modification on the target models, which hinders their practical application. To address this issue, we propose a…

机器学习 · 计算机科学 2023-12-27 Haonan Wang , Ziwei Wu , Jingrui He

Pac-Bayes bounds are among the most accurate generalization bounds for classifiers learned from independently and identically distributed (IID) data, and it is particularly so for margin classifiers: there have been recent contributions…

机器学习 · 计算机科学 2010-06-09 Liva Ralaivola , Marie Szafranski , Guillaume Stempfel

A novel general neural network (GNN) is proposed for two-class data mining in this study. In a GNN, each attribute in the dataset is treated as a node, with each pair of nodes being connected by an arc. The reliability is of each arc, which…

神经与进化计算 · 计算机科学 2019-10-24 Wei-Chang Yeh

This article explores the generalized analysis-of-variance or ANOVA dimensional decomposition (ADD) for multivariate functions of dependent random variables. Two notable properties, stemming from weakened annihilating conditions, reveal…

数值分析 · 数学 2014-08-05 Sharif Rahman

Reshef & Reshef recently published a paper in which they present a method called the Maximal Information Coefficient (MIC) that can detect all forms of statistical dependence between pairs of variables as sample size goes to infinity. While…

机器学习 · 统计学 2013-08-28 Alexander Luedtke , Linh Tran

Rigidity Percolation with g degrees of freedom per site is analyzed on randomly diluted Erdos-Renyi graphs with average connectivity gamma, in the presence of a field h. In the (gamma,h) plane, the rigid and flexible phases are separated by…

统计力学 · 物理学 2009-11-10 Cristian F. Moukarzel

Causal effect estimation relies on separating the variation in the outcome into parts due to the treatment and due to the confounders. To achieve this separation, practitioners often use external sources of randomness that only influence…

机器学习 · 计算机科学 2021-02-03 Aahlad Manas Puli , Rajesh Ranganath

Out-of-distribution (OOD) detection is paramount to ensuring the reliability and robustness of learning models in real-world applications. Existing post-hoc OOD detection methods detect OOD samples by leveraging their features and logits…

计算机视觉与模式识别 · 计算机科学 2025-11-05 Kun Zou , Yongheng Xu , Jianxing Yu , Yan Pan , Jian Yin , Hanjiang Lai

A common problem in numerous research areas, particularly in clinical trials, is to test whether the effect of an explanatory variable on an outcome variable is equivalent across different groups. In practice, these tests are frequently…

统计方法学 · 统计学 2024-05-03 Niklas Hagemann , Kathrin Möllenhoff

Over-parameterized models can perfectly learn various types of data distributions, however, generalization error is usually lower for real data in comparison to artificial data. This suggests that the properties of data distributions have…

机器学习 · 计算机科学 2022-07-28 Martin Briesch , Dominik Sobania , Franz Rothlauf

In causal inference, estimating the average treatment effect is a central objective, and in the context of competing risks data, this effect can be quantified by the cause-specific cumulative incidence function (CIF) difference. While…

统计方法学 · 统计学 2026-03-27 Yifei Tian , Ying Wu

Developing fast regression models (surrogate/metamodels) from DEM data is key for practical industrial application to allow real-time evaluations. However, benchmarking different models is often overlooked in particle technology for…

Adversarial robustness is essential for deploying neural networks in safety-critical applications, yet standard evaluation methods either require expensive adversarial attacks or report only a single aggregate score that obscures how…

机器学习 · 计算机科学 2026-04-15 Arya Shah , Kaveri Visavadiya , Manisha Padala