English
Related papers

Related papers: A testing-based approach to the discovery of diffe…

200 papers

Particle size is a key variable in understanding the behaviour of the particulate products that underpin much of our modern lives. Typically obtained from suspensions at rest, measuring the particle size under flowing conditions would…

Soft Condensed Matter · Physics 2021-07-22 James A Richards , Vincent A Martinez , Jochen Arlt

A representative model in integrative analysis of two high-dimensional correlated datasets is to decompose each data matrix into a low-rank common matrix generated by latent factors shared across datasets, a low-rank distinctive matrix…

Machine Learning · Statistics 2022-04-06 Hai Shu , Zhe Qu

Fields like public health, public policy, and social science often want to quantify the degree of dependence between variables whose relationships take on unknown functional forms. Typically, in fact, researchers in these fields are…

Statistics Theory · Mathematics 2019-12-10 Octavio César Mesner , Cosma Rohilla Shalizi

Causal discovery aims to learn causal relationships between variables from targeted data, making it a fundamental task in machine learning. However, causal discovery algorithms often rely on unverifiable causal assumptions, which are…

Machine Learning · Computer Science 2025-10-15 Huiyang Yi , Yanyan He , Duxin Chen , Mingyu Kang , He Wang , Wenwu Yu

Causal discovery studies the problem of mining causal relationships between variables from data, which is of primary interest in science. During the past decades, significant amount of progresses have been made toward this fundamental data…

Artificial Intelligence · Computer Science 2016-11-28 Kui Yu , Jiuyong Li , Lin Liu

Our aim is to detect mechanistic interaction between the effects of two causal factors on a binary response, as an aid to identifying situations where the effects are mediated by a common mechanism. We propose a formalization of mechanistic…

Methodology · Statistics 2015-06-23 Carlo Berzuini , A. Philip Dawid

Independence screening methods such as the two sample $t$-test and the marginal correlation based ranking are among the most widely used techniques for variable selection in ultrahigh dimensional data sets. In this short note, simple…

Methodology · Statistics 2020-11-17 Run Wang , Somak Dutta , Vivekananda Roy

Dynamic Causal Modelling (DCM) is the predominant method for inferring effective connectivity from neuroimaging data. In the 15 years since its introduction, the neural models and statistical routines in DCM have developed in parallel,…

Quantitative Methods · Quantitative Biology 2019-07-15 Peter Zeidman , Amirhossein Jafarian , Nadège Corbin , Mohamed L. Seghier , Adeel Razi , Cathy J. Price , Karl J. Friston

Computation of Mutual Information (MI) helps understand the amount of information shared between a pair of random variables. Automated feature selection techniques based on MI ranking are regularly used to extract information from sensitive…

Cryptography and Security · Computer Science 2020-09-24 Ankit Srivastava , Samira Pouyanfar , Joshua Allen , Ken Johnston , Qida Ma

Data stream classification methods demonstrate promising performance on a single data stream by exploring the cohesion in the data stream. However, multiple data streams that involve several correlated data streams are common in many…

Machine Learning · Computer Science 2019-08-19 Yingzhong Shi , Zhaohong Deng , Haoran Chen , Kup-Sze Choi , Shitong Wang

The focus of disentanglement approaches has been on identifying independent factors of variation in data. However, the causal variables underlying real-world observations are often not statistically independent. In this work, we bridge the…

Observability can determine which recorded variables of a given system are optimal for discriminating its different states. Quantifying observability requires knowledge of the equations governing the dynamics. These equations are often…

Adaptation and Self-Organizing Systems · Physics 2020-10-28 Christopher E. Gonzalez , Claudia Lainscsek , Terrence J. Sejnowski , Christophe Letellier

Co-clustering is a class of unsupervised data analysis techniques that extract the existing underlying dependency structure between the instances and variables of a data table as homogeneous blocks. Most of those techniques are limited to…

Machine Learning · Computer Science 2022-12-23 Aichetou Bouchareb , Marc Boullé , Fabrice Clérot , Fabrice Rossi

This paper introduces the correlation-of-divergency coefficient, c-delta, a custom statistical measure designed to quantify the similarity of internal divergence patterns between two groups of values. Unlike conventional correlation…

Methodology · Statistics 2026-03-10 Johan F. Hoorn

Correlations between factors of variation are prevalent in real-world data. Exploiting such correlations may increase predictive performance on noisy data; however, often correlations are not robust (e.g., they may change between domains,…

Machine Learning · Computer Science 2022-12-26 Christina M. Funke , Paul Vicol , Kuan-Chieh Wang , Matthias Kümmerer , Richard Zemel , Matthias Bethge

Dynamic Music Emotion Recognition (DMER) aims to predict the emotion of different moments in music, playing a crucial role in music information retrieval. The existing DMER methods struggle to capture long-term dependencies when dealing…

Sound · Computer Science 2024-12-30 Dengming Zhang , Weitao You , Ziheng Liu , Lingyun Sun , Pei Chen

Data Mining is the process of examining the information from different point of view and compressing it for the relevant data. This data can also be utilized to build the incomes. Data Mining is also known as Data or Knowledge Discovery.…

Databases · Computer Science 2016-10-17 Kratika Tyagi , Prof. Sanjeev Thakur

The discovery of causal relationships from high-dimensional data is a major open problem in bioinformatics. Machine learning and feature attribution models have shown great promise in this context but lack causal interpretation. Here, we…

Machine Learning · Computer Science 2023-04-26 Payam Dibaeinia , Saurabh Sinha

We introduce Compartmentalized Diffusion Models (CDM), a method to train different diffusion models (or prompts) on distinct data sources and arbitrarily compose them at inference time. The individual models can be trained in isolation, at…

Machine Learning · Computer Science 2024-10-15 Aditya Golatkar , Alessandro Achille , Ashwin Swaminathan , Stefano Soatto

We present the extention and application of a new unsupervised statistical learning technique--the Partition Decoupling Method--to gene expression data. Because it has the ability to reveal non-linear and non-convex geometries present in…

Quantitative Methods · Quantitative Biology 2015-09-24 Rosemary Braun , Gregory Leibon , Scott Pauls , Daniel Rockmore
‹ Prev 1 3 4 5 6 7 10 Next ›