中文
相关论文

相关论文: Notes on the H-measure of classifier performance

200 篇论文

In this exploratory note we ask the question of what a measure of performance for all tasks is like if we use a weighting of tasks based on a difficulty function. This difficulty function depends on the complexity of the (acceptable)…

人工智能 · 计算机科学 2015-03-27 Jose Hernandez-Orallo

This paper describes measures for evaluating the three determinants of how well a probabilistic classifier performs on a given test set. These determinants are the appropriateness, for the test set, of the results of (1) feature selection,…

cmp-lg · 计算机科学 2008-02-03 Rebecca Bruce , Janyce Wiebe , Ted Pedersen

H-measures and semiclassical (Wigner) measures were introduced in earlyn 1990s and since then they have found numerous applications in problems involving $\mathrm{L}^2$ weakly converging sequences. Although they are similar objects, neither…

偏微分方程分析 · 数学 2023-09-06 Nenad Antonić , Marko Erceg

The h-index has become a widely used metric for evaluating the productivity and citation impact of researchers. Introduced by physicist Jorge E. Hirsch in 2005, the h-index measures both the quantity (number of publications) and quality…

数字图书馆 · 计算机科学 2025-03-19 Ali Borji

We introduce the E-measure: a measure-like generalization of the E-value to a class of hypotheses. Unlike classical measures, E-measures are closed under infimums instead of addition. They arise from a compatibility axiom with logical…

统计理论 · 数学 2026-04-23 Nick W. Koning

The F-measure or F-score is one of the most commonly used single number measures in Information Retrieval, Natural Language Processing and Machine Learning, but it is based on a mistake, and the flawed assumptions render it unsuitable for…

信息检索 · 计算机科学 2019-09-13 David M. W. Powers

Fine-tuning of large pre-trained image and language models on small customized datasets has become increasingly popular for improved prediction and efficient use of limited resources. Fine-tuning requires identification of best models to…

机器学习 · 计算机科学 2023-05-29 Shibal Ibrahim , Natalia Ponomareva , Rahul Mazumder

The F-measure, also known as the F1-score, is widely used to assess the performance of classification algorithms. However, some researchers find it lacking in intuitive interpretation, questioning the appropriateness of combining two…

机器学习 · 计算机科学 2021-03-19 David J. Hand , Peter Christen , Nishadi Kirielle

Evaluating the performance of classifiers is critical in machine learning, particularly in high-stakes applications where the reliability of predictions can significantly impact decision-making. Traditional performance measures, such as…

机器学习 · 计算机科学 2024-12-19 Jesus S. Aguilar-Ruiz

Attribute weighting and differential weighting, two major mechanisms for computing context-dependent similarity or dissimilarity measures are studied and compared. A dissimilarity measure based on subset size in the context is proposed and…

人工智能 · 计算机科学 2013-04-05 Yizong Cheng

Classification is the task of predicting the class labels of objects based on the observation of their features. In contrast, quantification has been defined as the task of determining the prevalences of the different sorts of class labels…

机器学习 · 统计学 2016-08-15 Dirk Tasche

The h-index is a mainstream bibliometric indicator, since it is widely used in academia, research management and research policy. While its advantages have been highlighted, such as its simple calculation, it has also received widespread…

数字图书馆 · 计算机科学 2021-12-07 Grischa Fraumann , Ruediger Mutz

In over-identified models, misspecification -- the norm rather than exception -- fundamentally changes what estimators estimate. Different estimators imply different estimands rather than different efficiency for the same target. A review…

计量经济学 · 经济学 2026-02-23 Isaiah Andrews , Jiafeng Chen , Otavio Tecchio

Testing Machine Learning (ML) models and AI-Infused Applications (AIIAs), or systems that contain ML models, is highly challenging. In addition to the challenges of testing classical software, it is acceptable and expected that statistical…

机器学习 · 计算机科学 2022-10-28 George Kour , Marcel Zalmanovici , Orna Raz , Samuel Ackerman , Ateret Anaby-Tavor

Trust is a crucial factor affecting the adoption of machine learning (ML) models. Qualitative studies have revealed that end-users, particularly in the medical domain, need models that can express their uncertainty in decision-making…

机器学习 · 计算机科学 2023-04-21 Andrew Houston , Georgina Cosma

Estimating the effort and quality of a system is a critical step at the beginning of every software project. It is necessary to have reliable ways of calculating these measures, and, it is even better when the calculation can be done as…

软件工程 · 计算机科学 2012-07-11 Andreas Bollin , Abdollah Tabareh

In most machine learning applications, classification accuracy is not the primary metric of interest. Binary classifiers which face class imbalance are often evaluated by the $F_\beta$ score, area under the precision-recall curve, Precision…

机器学习 · 计算机科学 2018-03-02 Alan Mackey , Xiyang Luo , Elad Eban

In this paper we present the first steps towards hardening the science of measuring AI systems, by adopting metrology, the science of measurement and its application, and applying it to human (crowd) powered evaluations. We begin with the…

人工智能 · 计算机科学 2019-11-06 Chris Welty , Praveen Paritosh , Lora Aroyo

Measuring performance & quantifying a performance change are core evaluation techniques in programming language and systems research. Of 122 recent scientific papers, as many as 65 included experimental evaluation that quantified a…

统计方法学 · 统计学 2020-07-22 Tomas Kalibera , Richard Jones

In this paper, we introduce a new measure called Term_Class relevance to compute the relevancy of a term in classifying a document into a particular class. The proposed measure estimates the degree of relevance of a given term, in placing…

信息检索 · 计算机科学 2016-09-15 D S Guru , Mahamad Suhil