English
Related papers

Related papers: Detecting Correlated Gaussian Databases

200 papers

The problem of frequent pattern mining from non-temporal databases is studied extensively by various researchers working in areas of data mining, temporal databases and information retrieval. However, Conventional frequent pattern…

Databases · Computer Science 2016-04-19 Vangipuram Radhakrishna , P. V. Kumar , V. Janaki

Transposable data represents interactions among two sets of entities, and are typically represented as a matrix containing the known interaction values. Additional side information may consist of feature vectors specific to entities…

Machine Learning · Statistics 2014-04-29 Oluwasanmi Koyejo , Cheng Lee , Joydeep Ghosh

Change Detection (CD) enables the identification of alterations between images of the same area captured at different times. However, existing CD methods still struggle to address pseudo changes resulting from domain information differences…

Computer Vision and Pattern Recognition · Computer Science 2024-10-15 Yi Xiao , Bin Luo , Jun Liu , Xin Su , Wei Wang

For two correlated graphs which are independently sub-sampled from a common Erd\H{o}s-R\'enyi graph $\mathbf{G}(n, p)$, we wish to recover their \emph{latent} vertex matching from the observation of these two graphs \emph{without labels}.…

Statistics Theory · Mathematics 2022-05-31 Jian Ding , Hang Du

The problem of identifying change points in high-dimensional Gaussian graphical models (GGMs) in an online fashion is of interest, due to new applications in biology, economics and social sciences. The offline version of the problem, where…

Statistics Theory · Mathematics 2020-03-18 Hossein Keshavarz , George Michailidis

Fully coherent searches (over realistic ranges of parameter space and year-long observation times) for unknown sources of continuous gravitational waves are computationally prohibitive. Less expensive hierarchical searches divide the data…

General Relativity and Quantum Cosmology · Physics 2009-10-27 Holger J. Pletsch , Bruce Allen

We study the Nearest Neighbor Search (NNS) problem in a high-dimensional setting where data lies in a low-dimensional subspace and is corrupted by Gaussian noise. Specifically, we consider a semi-random model in which $n$ points from an…

Data Structures and Algorithms · Computer Science 2026-04-07 Ravindran Kannan , Kijun Shin , David Woodruff

We consider the problem of learning a graph modeling the statistical relations of the $d$ variables from a dataset with $n$ samples $X \in \mathbb{R}^{n \times d}$. Standard approaches amount to searching for a precision matrix $\Theta$…

Machine Learning · Statistics 2023-12-13 Titouan Vayer , Etienne Lasalle , Rémi Gribonval , Paulo Gonçalves

Understanding the relationships between different properties of data, such as whether a connectome or genome has information about disease status, is becoming increasingly important in modern biological datasets. While existing approaches…

Machine Learning · Statistics 2024-06-27 Joshua T. Vogelstein , Eric Bridgeford , Qing Wang , Carey E. Priebe , Mauro Maggioni , Cencheng Shen

Doubly intractable problems occur when both the likelihood and the posterior are available only in unnormalised form, with computationally intractable normalisation constants. Bayesian inference then typically requires direct approximation…

Anomaly detection using a network-based approach is one of the most efficient ways to identify abnormal events such as fraud, security breaches, and system faults in a variety of applied domains. While most of the earlier works address the…

Artificial Intelligence · Computer Science 2025-01-22 Hossein Rafieizadeh , Hadi Zare , Mohsen Ghassemi Parsa , Hadi Davardoust , Meshkat Shariat Bagheri

Gaussian graphical models emerge in a wide range of fields. They model the statistical relationships between variables as a graph, where an edge between two variables indicates conditional dependence. Unfortunately, well-established…

Machine Learning · Statistics 2024-01-19 Taulant Koka , Jasin Machkour , Michael Muma

Repairing inconsistent knowledge bases is a task that has been assessed, with great advances over several decades, from within the knowledge representation and reasoning and the database theory communities. As information becomes more…

Databases · Computer Science 2023-07-14 Sergio Abriola , Santiago Cifuentes , Nina Pardal , Edwin Pin

We study the problem of learning the topology of a directed Gaussian Graphical Model under the equal-variance assumption, where the graph has $n$ nodes and maximum in-degree $d$. Prior work has established that $O(d \log n)$ samples are…

Machine Learning · Computer Science 2025-11-11 Constantinos Daskalakis , Vardis Kandiros , Rui Yao

We consider the problem of aligning a pair of databases with correlated entries. We introduce a new measure of correlation in a joint distribution that we call cycle mutual information. This measure has operational significance: it…

Information Theory · Computer Science 2018-05-11 Daniel Cullina , Prateek Mittal , Negar Kiyavash

We tackle distributed detection of a non-cooperative target with a Wireless Sensor Network (WSN). When the target is present, sensors observe an (unknown) deterministic signal with attenuation depending on the distance between the sensor…

Information Theory · Computer Science 2017-04-26 D. Ciuonzo , P. Salvo Rossi , P. Willett

Recovery of the causal structure of dynamic networks from noisy measurements has long been a problem of interest across many areas of science and engineering. Many algorithms have been proposed, but there is little work that compares the…

Information Theory · Computer Science 2025-06-10 Xiaohan Kang , Bruce Hajek

Consider a supervised dataset $D=[A\mid \textbf{b}]$, where $\textbf{b}$ is the outcome column, rows of $D$ correspond to observations, and columns of $A$ are the features of the dataset. A central problem in machine learning and pattern…

Machine Learning · Computer Science 2019-02-27 Javad Rahimipour Anaraki , Hamid Usefi

The problem of detecting edge correlation between two Erd\H{o}s-R\'enyi random graphs on $n$ unlabeled nodes can be formulated as a hypothesis testing problem: under the null hypothesis, the two graphs are sampled independently; under the…

Probability · Mathematics 2022-05-31 Jian Ding , Hang Du

Recently, different works proposed a new way to mine patterns in databases with pathological size. For example, experiments in genome biology usually provide databases with thousands of attributes (genes) but only tens of objects…

Machine Learning · Computer Science 2009-02-10 Baptiste Jeudy , François Rioult