中文
相关论文

相关论文: Assessing Uncertainty in Similarity Scoring: Perfo…

200 篇论文

Face recognition is a widely used authentication technology in practice, where robustness is required. It is thus essential to have an efficient and easy-to-use method for evaluating the robustness of (possibly third-party) trained face…

软件工程 · 计算机科学 2025-05-01 Ruihan Zhang , Jun Sun

Despite being widely used, face recognition models suffer from bias: the probability of a false positive (incorrect face match) strongly depends on sensitive attributes such as the ethnicity of the face. As a result, these models can…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Tiago Salvador , Stephanie Cairns , Vikram Voleti , Noah Marshall , Adam Oberman

Measuring the accuracy of face recognition (FR) systems is essential for improving performance and ensuring responsible use. Accuracy is typically estimated using large annotated datasets, which are costly and difficult to obtain. We…

计算机视觉与模式识别 · 计算机科学 2025-02-24 Manuel Knott , Ignacio Serna , Ethan Mann , Pietro Perona

While the area under the ROC curve is perhaps the most common measure that is used to rank the relative performance of different binary classifiers, longstanding field folklore has noted that it can be a measure that ill-captures the…

机器学习 · 计算机科学 2024-12-19 Christopher Ratigan , Lenore Cowen

Many recent news headlines have labeled face recognition technology as biased or racist. We report on a methodical investigation into differences in face recognition accuracy between African-American and Caucasian image cohorts of the MORPH…

计算机视觉与模式识别 · 计算机科学 2019-05-09 KS Krishnapriya , Kushal Vangara , Michael C. King , Vitor Albiero , Kevin Bowyer

With growing applications of Machine Learning (ML) techniques in the real world, it is highly important to ensure that these models work in an equitable manner. One main step in ensuring fairness is to effectively measure fairness, and to…

机器学习 · 计算机科学 2024-06-21 Abdalwahab Almajed , Maryam Tabar , Peyman Najafirad

The Receiver Operating Characteristic (ROC) curve stands as a cornerstone in assessing the efficacy of biomarkers for disease diagnosis. Beyond merely evaluating performance, it provides with an optimal cutoff for biomarker values, crucial…

统计方法学 · 统计学 2025-04-29 Soutik Ghosal

There are strong incentives to build models that demonstrate outstanding predictive performance on various datasets and benchmarks. We believe these incentives risk a narrow focus on models and on the performance metrics used to evaluate…

机器学习 · 计算机科学 2022-06-07 David Lovell , Dimity Miller , Jaiden Capra , Andrew Bradley

The proper use of model evaluation metrics is important for model evaluation and model selection in binary classification tasks. This study investigates how consistent different metrics are at evaluating models across data of different…

机器学习 · 统计学 2024-12-17 Jing Li

Fairness has been a critical issue that affects the adoption of deep learning models in real practice. To improve model fairness, many existing methods have been proposed and evaluated to be effective in their own contexts. However, there…

机器学习 · 计算机科学 2024-03-26 Junjie Yang , Jiajun Jiang , Zeyu Sun , Junjie Chen

Before deploying a black-box model in high-stakes problems, it is important to evaluate the model's performance on sensitive subpopulations. For example, in a recidivism prediction task, we may wish to identify demographic groups for which…

统计方法学 · 统计学 2023-06-09 John J. Cherian , Emmanuel J. Candès

The receiver operating characteristic (ROC) curve is the most popular tool used to evaluate the discriminatory capability of diagnostic tests/biomarkers measured on a continuous scale when distinguishing between two alternative disease…

统计方法学 · 统计学 2021-03-22 Maria Xose Rodriguez-Alvarez , Vanda Inacio

Throughout science and technology, receiver operating characteristic (ROC) curves and associated area under the curve (AUC) measures constitute powerful tools for assessing the predictive abilities of features, markers and tests in binary…

机器学习 · 统计学 2021-06-25 Tilmann Gneiting , Eva-Maria Walz

The urging societal demand for fair AI systems has put pressure on the research community to develop predictive models that are not only globally accurate but also meet new fairness criteria, reflecting the lack of disparate mistreatment…

计算机视觉与模式识别 · 计算机科学 2025-04-29 Jean-Rémy Conti , Stéphan Clémençon

In this paper we compare two regression curves by measuring their difference by the area between the two curves, represented by their $L^1$-distance. We develop asymptotic confidence intervals for this measure and statistical tests to…

统计理论 · 数学 2023-02-03 Patrick Bastian , Holger Dette , Lukas Koletzko , Kathrin Möllenhoff

The development of face recognition algorithms by academic and commercial organizations is growing rapidly due to the onset of deep learning and the widespread availability of training data. Though tests of face recognition algorithm…

计算机视觉与模式识别 · 计算机科学 2022-03-11 John J. Howard , Eli J. Laird , Yevgeniy B. Sirotin , Rebecca E. Rubin , Jerry L. Tipton , Arun R. Vemury

The optimal receiver operating characteristic (ROC) curve, giving the maximum probability of detection as a function of the probability of false alarm, is a key information-theoretic indicator of the difficulty of a binary hypothesis…

信息论 · 计算机科学 2025-06-10 Bruce Hajek , Xiaohan Kang

The prevalence of algorithmic bias in Machine Learning (ML)-driven approaches has inspired growing research on measuring and mitigating bias in the ML domain. Accordingly, prior research studied how to measure fairness in regression which…

机器学习 · 计算机科学 2025-08-21 Abdalwahab Almajed , Maryam Tabar , Peyman Najafirad

Whilst the size and complexity of ML models have rapidly and significantly increased over the past decade, the methods for assessing their performance have not kept pace. In particular, among the many potential performance metrics, the ML…

机器学习 · 计算机科学 2023-12-29 Michael Roberts , Alon Hazan , Sören Dittmer , James H. F. Rudd , Carola-Bibiane Schönlieb

Many fields use the ROC curve and the PR curve as standard evaluations of binary classification methods. Analysis of ROC and PR, however, often gives misleading and inflated performance evaluations, especially with an imbalanced ground…

机器学习 · 统计学 2020-06-23 Chang Cao , Davide Chicco , Michael M. Hoffman