English
Related papers

Related papers: Standardizing Interestingness Measures for Associa…

200 papers

Measuring interdisciplinarity is a pertinent but challenging issue in quantitative studies of science. There seems to be a consensus in the literature that the concept of interdisciplinarity is multifaceted and ambiguous. Unsurprisingly,…

Digital Libraries · Computer Science 2019-12-11 Qi Wang , Jesper Wiborg Schneider

Measures of irrationality are a numerical way of quantifying how far a given variety is from being rational (or rationally connected, uniruled, etc.). In the last two decades, there has been renewed interest in the study of these…

Algebraic Geometry · Mathematics 2025-09-05 Nathan Chen , Olivier Martin

Several rules for social choice are examined from a unifying point of view that looks at them as procedures for revising a system of degrees of belief in accordance with certain specified logical constraints. Belief is here a social…

Artificial Intelligence · Computer Science 2015-05-06 Rosa Camps , Xavier Mora , Laia Saumell

In this study, we address the challenge of measuring the ability of a recommender system to make surprising recommendations. Although current evaluation methods make it possible to determine if two algorithms can make recommendations with a…

Information Retrieval · Computer Science 2018-07-12 Andre Paulino de Lima , Sarajane Marques Peres

The solution to fine tuning is one of the principal motivations for Beyond the Standard Model (BSM) Studies. However constraints on new physics indicate that many of these BSM models are also fine tuned (although to a much lesser extent).…

High Energy Physics - Phenomenology · Physics 2008-11-26 Peter Athron , D. J. Miller

Decision making is challenging when there is more than one criterion to consider. In such cases, it is common to assign a goodness score to each item as a weighted sum of its attribute values and rank them accordingly. Clearly, the ranking…

Databases · Computer Science 2018-12-20 Abolfazl Asudeh , H. V. Jagadish , Gerome Miklau , Julia Stoyanovich

Despite the clear performance benefits of data augmentations, little is known about why they are so effective. In this paper, we disentangle several key mechanisms through which data augmentations operate. Establishing an exchange rate…

Machine Learning · Computer Science 2023-04-03 Jonas Geiping , Micah Goldblum , Gowthami Somepalli , Ravid Shwartz-Ziv , Tom Goldstein , Andrew Gordon Wilson

Individual fairness is an intuitive definition of algorithmic fairness that addresses some of the drawbacks of group fairness. Despite its benefits, it depends on a task specific fair metric that encodes our intuition of what is fair and…

Machine Learning · Statistics 2020-06-23 Debarghya Mukherjee , Mikhail Yurochkin , Moulinath Banerjee , Yuekai Sun

The distance standard deviation, which arises in distance correlation analysis of multivariate data, is studied as a measure of spread. The asymptotic distribution of the empirical distance standard deviation is derived under the assumption…

Statistics Theory · Mathematics 2019-12-12 Dominic Edelmann , Donald Richards , Daniel Vogel

Applied researchers often claim that the risk difference is more heterogeneous than the relative risk and the odds ratio. Some also argue that there are theoretical grounds for why this claim is true. In this note, we point out that these…

Applications · Statistics 2022-10-12 Linbo Wang

Given a set of people and a set of events they attend, we address the problem of measuring connectedness or tie strength between each pair of persons given that attendance at mutual events gives an implicit social network between people. We…

Social and Information Networks · Computer Science 2011-12-14 Mangesh Gupte , Tina Eliassi-Rad

We identify the task of measuring data to quantitatively characterize the composition of machine learning data and datasets. Similar to an object's height, width, and volume, data measurements quantify different attributes of data along…

Quantifying the similarity between two mathematical structures or datasets constitutes a particularly interesting and useful operation in several theoretical and applied problems. Aimed at this specific objective, the Jaccard index has been…

Machine Learning · Computer Science 2021-11-19 Luciano da F. Costa

Score matching is a recently developed parameter learning method that is particularly effective to complicated high dimensional density models with intractable partition functions. In this paper, we study two issues that have not been…

Machine Learning · Computer Science 2012-05-14 Siwei Lyu

Some existing notions of redundancy among association rules allow for a logical-style characterization and lead to irredundant bases of absolutely minimum size. One can push the intuition of redundancy further and find an intuitive notion…

Databases · Computer Science 2011-03-25 José L. Balcázar

In this paper we consider the axiomatic characterization of information and certainty measures in a unified way. We present the general axiomatic system which captures the common properties of a large number of the measures previously…

Information Theory · Computer Science 2015-06-17 Velimir M. Ilic , Miomir S. Stankovic

Mutual information is commonly used as a measure of similarity between competing labelings of a given set of objects, for example to quantify performance in classification and community detection tasks. As argued recently, however, the…

Social and Information Networks · Computer Science 2025-07-17 Maximilian Jerdee , Alec Kirkley , M. E. J. Newman

The solution to fine tuning is one of the principal motivations for supersymmetry. However constraints on the parameter space of the Minimal Supersymmetric Standard Model (MSSM) suggest it may also require fine tuning (although to a much…

High Energy Physics - Phenomenology · Physics 2007-10-15 Peter Athron , D. J. Miller

Experimental datasets are growing rapidly in size, scope, and detail, but the value of these datasets is limited by unwanted measurement noise. It is therefore tempting to apply analysis techniques that attempt to reduce noise and enhance…

Applications · Statistics 2022-07-12 Kendrick Kay

Similarity join, which can find similar objects (e.g., products, names, addresses) across different sources, is powerful in dealing with variety in big data, especially web data. Threshold-driven similarity join, which has been extensively…

Databases · Computer Science 2017-07-13 Chuancong Gao , Jiannan Wang , Jian Pei , Rui Li , Yi Chang