English
Related papers

Related papers: Mathematical Foundations of Data Cohesion

200 papers

A linear transformation f(S) of configurational entropy with length scale dependent coefficients as a measure of spatial inhomogeneity is considered. When a final pattern is formed with periodically repeated initial arrangement of point…

Statistical Mechanics · Physics 2009-10-31 Z. Garncarek , R. Piasecki

Aggregation functions are generally defined and used to combine several numerical values into a single one, so that the final result of the aggregation takes into account all the individual values in a given manner. Such functions are…

Statistics Theory · Mathematics 2009-06-22 Jean-Luc Marichal

Data represented by probability measures arise as empirical distributions, posterior distributions, and feature-based representations of complex objects. We study heterogeneity in a population of probability measures through the expected…

Methodology · Statistics 2026-03-17 Kisung You

Given a decision process based on the approximate probability density function returned by a data assimilation algorithm, an interaction level between the decision making level and the data assimilation level is designed to incorporate the…

Computation · Statistics 2015-03-19 Gabriel Terejanu , Puneet Singla , Tarunraj Singh , Peter D. Scott

The principle of similarity, or homophily, is often used to explain patterns observed in complex networks such as transitivity and the abundance of triangles (3-cycles). However, many phenomena from division of labor to protein-protein…

Physics and Society · Physics 2022-10-12 Szymon Talaga , Andrzej Nowak

Feature selection can facilitate the learning of mixtures of discrete random variables as they arise, e.g. in crowdsourcing tasks. Intuitively, not all workers are equally reliable but, if the less reliable ones could be eliminated, then…

Machine Learning · Statistics 2017-11-28 Vincent Zhao , Steven W. Zucker

Many real-life data are described by categorical attributes without a pre-classification. A common data mining method used to extract information from this type of data is clustering. This method group together the samples from the data…

Machine Learning · Computer Science 2014-07-30 Fabricio Olivetti de França

The traditional approach to the quantitative study of segregation is to employ indices that are selected by ``desirable properties''. Here, we detail how information theory underpins entropy-based indices and demonstrate how desirable…

Physics and Society · Physics 2022-12-15 Boris Barron , Yunus A. Kinkhabwala , Chriss Hess , Matthew Hall , Itai Cohen , Tomás A. Arias

The difficulties of detecting association, measuring correlation, and establishing cause and effect have fascinated mankind since time immemorial. Democritus, the Greek philosopher, underscored well the importance and the difficulty of…

Other Statistics · Statistics 2017-09-20 Donald St. P. Richards

We present Collaborative Trees, a novel tree model designed for regression prediction, along with its bagging version, which aims to analyze complex statistical associations between features and uncover potential patterns inherent in the…

Methodology · Statistics 2024-05-21 Chien-Ming Chi

Coherence is a defining property of quantum theory that accounts for quantum advantage in many quantum information tasks. Although many coherence quantifiers have been introduced in various contexts, the lack of efficient methods to…

Quantum Physics · Physics 2023-01-02 Sun Liang Liang , Yu Sixia

Community detection is a challenging and relevant problem in various disciplines of science and engineering like power systems, gene-regulatory networks, social networks, financial networks, astronomy etc. Furthermore, in many of these…

Systems and Control · Electrical Eng. & Systems 2022-04-06 Subhrajit Sinha

The increasing practice of engaging crowds, where organizations use IT to connect with dispersed individuals for explicit resource creation purposes, has precipitated the need to measure the precise processes and benefits of these…

Computers and Society · Computer Science 2017-02-15 J. Prpic , P. , Shukla

Quantum entanglement is a useful resource for implementing communication tasks. However, for the resource to be useful in practice, it needs to be accessible by parties with bounded computational resources. Computational entanglement…

Quantum Physics · Physics 2025-09-29 Ilia Ryzov , Faedi Loulidi , David Elkouss

Humans cluster in social groups where they discuss their shared past, problems, and potential solutions; they learn collectively when they repeat activities; they establish social norms; they synchronize when they sing or dance together;…

Social and Information Networks · Computer Science 2024-10-07 Jeroen Bruggeman

We survey the emerging area of compression-based, parameter-free, similarity distance measures useful in data-mining, pattern recognition, learning and automatic semantics extraction. Given a family of distances on a set of objects, a…

Computer Vision and Pattern Recognition · Computer Science 2007-05-23 Rudi Cilibrasi , Paul Vitanyi

Heterogeneous datasets emerge in various machine learning and optimization applications that feature different input sources, types or formats. Most models or methods do not natively tackle heterogeneity. Hence, such datasets are often…

Machine Learning · Statistics 2025-08-25 Edward Hallé-Hannan , Charles Audet , Youssef Diouane , Sébastien Le Digabel , Paul Saves

This work focuses on the characterization of the central tendency of a sample of compositional data. It provides new results about theoretical properties of means and covariance functions for compositional data, with an axiomatic…

Methodology · Statistics 2017-10-24 Denis Allard , Thierry Marchant

This paper presents a distance function between sets based on an average of distances between their elements. The distance function is a metric if the sets are non-empty finite subsets of a metric space. It can be applied to produce various…

Metric Geometry · Mathematics 2011-09-13 Osamu Fujita

Real-world data typically contain a large number of features that are often heterogeneous in nature, relevance, and also units of measure. When assessing the similarity between data points, one can build various distance measures using…

Machine Learning · Statistics 2022-05-27 Aldo Glielmo , Claudio Zeni , Bingqing Cheng , Gabor Csanyi , Alessandro Laio
‹ Prev 1 4 5 6 7 8 10 Next ›