English
Related papers

Related papers: Controlling False Positives in Association Rule Mi…

200 papers

The performance of active learning algorithms can be improved in two ways. The often used and intuitive way is by reducing the overall error rate within the test set. The second way is to ensure that correct predictions are not forgotten…

Machine Learning · Computer Science 2024-11-19 Ryan Benkert , Mohit Prabhushankar , Ghassan AlRegib

This paper presents an approach of using methods of process mining and rule-based artificial intelligence to analyze and understand study paths of students based on campus management system data and study program models. Process mining…

In research on eye movements in reading, it is common to analyze a number of canonical dependent measures to study how the effects of a manipulation unfold over time. Although this gives rise to the well-known multiple comparisons problem,…

Applications · Statistics 2016-10-11 Titus von der Malsburg , Bernhard Angele

Science and technology have a growing need for effective mechanisms that ensure reliable, controlled performance from black-box machine learning algorithms. These performance guarantees should ideally hold conditionally on the input-that is…

Machine Learning · Computer Science 2025-03-28 Vincent Blot , Anastasios N Angelopoulos , Michael I Jordan , Nicolas J-B Brunel

Scientific fraud is an increasingly vexing problem. Many current programs for fraud detection focus on image manipulation, while techniques for detection based on anomalous patterns that may be discoverable in the underlying numerical data…

Quantitative Methods · Quantitative Biology 2013-11-22 Joel H Pitt , Helene Z Hill

Granular association rule is a new approach to reveal patterns hide in many-to-many relationships of relational databases. Different types of data such as nominal, numeric and multi-valued ones should be dealt with in the process of rule…

Information Retrieval · Computer Science 2013-05-08 Fan Min , William Zhu

A new online multiple testing procedure is described in the context of anomaly detection, which controls the False Discovery Rate (FDR). An accurate anomaly detector must control the false positive rate at a prescribed level while keeping…

Methodology · Statistics 2024-12-17 Etienne Krönert , Alain Célisse , Dalila Hattab

The business processes of organizations may deviate from normal control flow due to disruptive anomalies, including unknown, skipped, and wrongly-ordered activities. To identify these control-flow anomalies, process mining can check…

Machine Learning · Computer Science 2025-02-17 Francesco Vitale , Marco Pegoraro , Wil M. P. van der Aalst , Nicola Mazzocca

When testing multiple hypothesis in a survey --e.g. many different source locations, template waveforms, and so on-- the final result consists in a set of confidence intervals, each one at a desired confidence level. But the probability…

General Relativity and Quantum Cosmology · Physics 2009-11-11 L. Baggio , G. A. Prodi

Popularity bias is a persistent issue associated with recommendation systems, posing challenges to both fairness and efficiency. Existing literature widely acknowledges that reducing popularity bias often requires sacrificing recommendation…

Information Retrieval · Computer Science 2023-06-05 Bin Liu , Erjia Chen , Bang Wang

Online testing procedures assume that hypotheses are observed in sequence, and allow the significance thresholds for upcoming tests to depend on the test statistics observed so far. Some of the most popular online methods include alpha…

Methodology · Statistics 2022-02-11 Aaron Fisher

Distantly supervised relation extraction (RE) automatically aligns unstructured text with relation instances in a knowledge base (KB). Due to the incompleteness of current KBs, sentences implying certain relations may be annotated as N/A…

Computation and Language · Computer Science 2021-09-07 Kailong Hao , Botao Yu , Wei Hu

Scientists have been interested in estimating causal peer effects to understand how people's behaviors are affected by their network peers. However, it is well known that identification and estimation of causal peer effects are challenging…

Methodology · Statistics 2021-09-07 Naoki Egami , Eric J. Tchetgen Tchetgen

In several different applications, including data transformation and entity resolution, rules are used to capture aspects of knowledge about the application at hand. Often, a large set of such rules is generated automatically or…

Artificial Intelligence · Computer Science 2020-11-03 Phokion G. Kolaitis , Lucian Popa , Kun Qian

An alternative formulation for the controllability problem of single input linear positive systems is presented. Driven by many industrial applications, this formulations focuses on the case where the region of interest is only a subset of…

Optimization and Control · Mathematics 2017-04-25 Yashar Zeinaly , Jan H. van Schuppen , Bart De Schutter

Anomaly detection when observing a large number of data streams is essential in a variety of applications, ranging from epidemiological studies to monitoring of complex systems. High-dimensional scenarios are usually tackled with…

Methodology · Statistics 2025-12-18 Ivo V. Stoepker , Rui M. Castro , Ery Arias-Castro , Edwin van den Heuvel

Large-scale hypothesis testing is central to modern science, where controlling the False Discovery Rate (FDR) has become the standard approach to managing false positives across many simultaneous tests. Hypotheses rarely exist in isolation;…

Methodology · Statistics 2026-05-19 Binyamin Perets , Shie Mannor

Machine learning operates at the intersection of statistics and computer science. This raises the question as to its underlying methodology. While much emphasis has been put on the close link between the process of learning from data and…

Machine Learning · Computer Science 2022-08-10 Oliver Buchholz , Eric Raidl

Multiple hypothesis testing, a situation when we wish to consider many hypotheses, is a core problem in statistical inference that arises in almost every scientific field. In this setting, controlling the false discovery rate (FDR), which…

Statistics Theory · Mathematics 2019-03-19 Shiyun Chen , Shiva Kasiviswanathan

Numerical Association Rule Mining is a popular variant of Association Rule Mining, where numerical attributes are handled without discretization. This means that the algorithms for dealing with this problem can operate directly, not only…

Neural and Evolutionary Computing · Computer Science 2020-10-30 Iztok Fister , Iztok Fister
‹ Prev 1 8 9 10 Next ›