中文
相关论文

相关论文: Standardizing Interestingness Measures for Associa…

200 篇论文

Though data augmentation has become a standard component of deep neural network training, the underlying mechanism behind the effectiveness of these techniques remains poorly understood. In practice, augmentation policies are often chosen…

机器学习 · 计算机科学 2020-06-08 Raphael Gontijo-Lopes , Sylvia J. Smullin , Ekin D. Cubuk , Ethan Dyer

We explore the possibility of using machine learning to identify interesting mathematical structures by using certain quantities that serve as fingerprints. In particular, we extract features from integer sequences using two empirical laws:…

机器学习 · 计算机科学 2018-09-11 Chai Wah Wu

The concept of complexity appears in virtually all areas of knowledge. Its intuitive meaning shares similarities across fields, but disagreements between its details hinders a general definition, leading to a plethora of proposed…

统计力学 · 物理学 2023-10-04 Roberto C. Alamino

The polarization measure is the probability that among 3 individuals chosen at random from a finite population exactly 2 come from the same class. This index is maximum at the midpoints of the edges of the probability simplex. We compute…

统计理论 · 数学 2015-04-22 Giovanni Pistone , Maria Piera Rogantin

Discovering relevant patterns for a particular user remains a challenging tasks in data mining. Several approaches have been proposed to learn user-specific pattern ranking functions. These approaches generalize well, but at the expense of…

人工智能 · 计算机科学 2022-03-08 Nassim Belmecheri , Noureddine Aribi , Nadjib Lazaar , Yahia Lebbah , Samir Loudni

Association Rule Mining is a machine learning method for discovering the interesting relations between the attributes in a huge transaction database. Typically, algorithms for Association Rule Mining generate a huge number of association…

神经与进化计算 · 计算机科学 2021-04-19 Iztok Fister , Iztok Fister

In the task of information retrieval the term relevance is taken to mean formal conformity of a document given by the retrieval system to user's information query. As a rule, the documents found by the retrieval system should be submitted…

计算与语言 · 计算机科学 2007-10-02 S. Braichevsky , D. Lande , A. Snarskii

Similarity is a fundamental measure in network analyses and machine learning algorithms, with wide applications ranging from personalized recommendation to socio-economic dynamics. We argue that an effective similarity measurement should…

物理与社会 · 物理学 2015-12-07 Jian-Guo Liu , Lei Hou , Xue Pan , Qiang Guo , Tao Zhou

A novel measure, quantumness of correlations is introduced here for bipartite states, by incorporating the required measurement scheme crucial in defining any such quantity. Quantumness coincides with the previously proposed measures in…

量子物理 · 物理学 2008-04-20 A. R. Usha Devi , A. K. Rajagopal

Statistical significance testing of differences in values of metrics like recall, precision and balanced F-score is a necessary part of empirical natural language processing. Unfortunately, we find in a set of experiments that many commonly…

计算与语言 · 计算机科学 2007-05-23 Alexander Yeh

Association rules express implication formed relations among attributes in databases of itemsets. The apriori algorithm is presented, the basis for most association rule mining algorithms. It works by pruning away rules that need not be…

数据库 · 计算机科学 2019-07-24 Niels Mündler

The basic idea of importance sampling is to use independent samples from a proposal measure in order to approximate expectations with respect to a target measure. It is key to understand how many samples are required in order to guarantee…

统计计算 · 统计学 2017-01-17 S. Agapiou , O. Papaspiliopoulos , D. Sanz-Alonso , A. M. Stuart

One often finds in the literature connections between measures of fairness and measures of feature importance employed to interpret trained classifiers. However, there seems to be no study that compares fairness measures and feature…

机器学习 · 计算机科学 2019-10-15 Juliana Cesaro , Fabio G. Cozman

Statistical practice does not automatically follow methodological innovation. Regularization methods, widely advocated to reduce overfitting and stabilize inference, are readily available in modern software, but are not consistently used by…

Statistical extreme value theory is concerned with the use of asymptotically motivated models to describe the extreme values of a process. A number of commonly used models are valid for observed data that exceed some high threshold.…

统计方法学 · 统计学 2014-12-10 J. Lee , Y. Fan , S. A. Sisson

We introduce a quantitative method to compare arbitrary pairs of graph centrality measures, based on the ordering of vertices induced by them. The proposed method is conceptually simple, mathematically elegant, and allows for a quantitative…

社会与信息网络 · 计算机科学 2026-01-26 G. Exarchakos , R. van der Hofstad , O. Nagy , M. Pandey

A ranking is an ordered sequence of items, in which an item with higher ranking score is more preferred than the items with lower ranking scores. In many information systems, rankings are widely used to represent the preferences over a set…

人工智能 · 计算机科学 2017-09-22 Zhiwei Lin , Yi Li , Xiaolian Guo

Einstein's Equivalence Principle is used with the electromagnetic spectrum to translate meters and seconds into radians and seconds. Based on a unique geometric relationship, a new transformation of velocities and a changed Lorentz…

综合物理 · 物理学 2007-05-23 Russell Clark Eskew

In discriminating between objects from different classes, the more separable these classes are the less computationally expensive and complex a classifier can be used. One thus seeks a measure that can quickly capture this separability…

统计方法学 · 统计学 2008-12-08 Linda Mthembu , Tshilidzi Marwala

A property, or statistical functional, is said to be elicitable if it minimizes expected loss for some loss function. The study of which properties are elicitable sheds light on the capabilities and limitations of point estimation and…

机器学习 · 计算机科学 2020-08-31 Rafael Frongillo , Ian A. Kash