English
Related papers

Related papers: Field Formulation of Parzen Data Analysis

200 papers

In cluster analysis, a common first step is to scale the data aiming to better partition them into clusters. Even though many different techniques have throughout many years been introduced to this end, it is probably fair to say that the…

Machine Learning · Computer Science 2023-05-30 Eduardo J. Aguilar , Valmir C. Barbosa

The nonparametric formulation of density-based clustering, known as modal clustering, draws a correspondence between groups and the attraction domains of the modes of the density function underlying the data. Its probabilistic foundation…

Methodology · Statistics 2020-10-27 Federico Ferraccioli , Giovanna Menardi

Topological data analysis (TDA) is an emerging mathematical concept for characterizing shapes in complex data. In TDA, persistence diagrams are widely recognized as a useful descriptor of data, and can distinguish robust and noisy…

Algebraic Topology · Mathematics 2016-04-27 Genki Kusano , Kenji Fukumizu , Yasuaki Hiraoka

Face clustering plays an essential role in exploiting massive unlabeled face data. Recently, graph-based face clustering methods are getting popular for their satisfying performances. However, they usually suffer from excessive memory…

Computer Vision and Pattern Recognition · Computer Science 2022-05-27 Junfu Liu , Di Qiu , Pengfei Yan , Xiaolin Wei

We consider the problem of analyzing the heterogeneity of clustering distributions for multiple groups of observed data, each of which is indexed by a covariate value, and inferring global clusters arising from observations aggregated over…

Methodology · Statistics 2012-12-06 XuanLong Nguyen

DBSCAN is one of the most important non-parametric unsupervised data analysis tools. By applying DBSCAN to a dataset, two key analytical results can be obtained: (1) clustering data points based on density distribution and (2) identifying…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Yongyu Wang

Quantum Clustering is a powerful method to detect clusters in data with mixed density. However, it is very sensitive to a length parameter that is inherent to the Schr\"odinger equation. In addition, linking data points into clusters…

Maxima of the linear density field form a point process that can be used to understand the spatial distribution of virialized halos that collapsed from initially overdense regions. However, owing to the peak constraint, clustering…

Cosmology and Nongalactic Astrophysics · Physics 2013-05-30 Vincent Desjacques

Clustering multi-dimensional points is a fundamental task in many fields, and density-based clustering supports many applications as it can discover clusters of arbitrary shapes. This paper addresses the problem of Density-Peaks Clustering…

Databases · Computer Science 2022-12-01 Daichi Amagata , Takahiro Hara

The clustering problem, and more generally, latent factor discovery --or latent space inference-- is formulated in terms of the Wasserstein barycenter problem from optimal transport. The objective proposed is the maximization of the…

Optimization and Control · Mathematics 2026-02-18 Hongkang Yang , Esteban G. Tabak

For numerous reasons there raises a need for dimension reduction that preserves certain characteristics of data. In this work we focus on data coming from a mixture of Gaussian distributions and we propose a method that preserves…

Statistics Theory · Mathematics 2014-07-30 Ewa Nowakowska , Jacek Koronacki , Stan Lipovetsky

The kernel exponential family is a rich class of distributions, which can be fit efficiently and with statistical guarantees by score matching. Being required to choose a priori a simple kernel such as the Gaussian, however, limits its…

Machine Learning · Statistics 2021-01-15 Li Wenliang , Danica J. Sutherland , Heiko Strathmann , Arthur Gretton

We derive necessary density conditions for sampling and for interpolation in general reproducing kernel Hilbert spaces satisfying some natural conditions on the geometry of the space and the reproducing kernel. If the volume of shells is…

Functional Analysis · Mathematics 2018-04-03 Hartmut Führ , Karlheinz Gröchenig , Antti Haimi , Andreas Klotz , José Luis Romero

The curse of dimensionality is a common phenomenon which affects analysis of datasets characterized by large numbers of variables associated with each point. Problematic scenarios of this type frequently arise in classification algorithms…

Probability · Mathematics 2015-08-11 Benjamin Thirey , Randal Hickman

Model-based clustering is widely-used in a variety of application areas. However, fundamental concerns remain about robustness. In particular, results can be sensitive to the choice of kernel representing the within-cluster data density.…

Machine Learning · Statistics 2019-06-27 Leo L Duan , David B Dunson

Crowd counting is a fundamental problem in crowd analysis which is typically accomplished by estimating a crowd density map and summing over the density values. However, this approach suffers from background noise accumulation and loss of…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Yasiru Ranasinghe , Nithin Gopalakrishnan Nair , Wele Gedara Chaminda Bandara , Vishal M. Patel

It is a consensus in signal processing that the Gaussian kernel and its partial derivatives enable the development of robust algorithms for feature detection. Fourier analysis and convolution theory have central role in such development. In…

Computer Vision and Pattern Recognition · Computer Science 2016-05-03 Paulo Sérgio Silva Rodrigues , Gilson Antonio Giraldi

One way of recovering information about the initial conditions of the Universe is by measuring features of the cosmological density field which are preserved during gravitational evolution and galaxy formation. In this paper we study the…

Astrophysics · Physics 2016-08-30 Rupert A. C. Croft , Enrique Gaztanaga

Semi-supervised clustering is the task of clustering data points into clusters where only a fraction of the points are labelled. The true number of clusters in the data is often unknown and most models require this parameter as an input.…

Machine Learning · Computer Science 2013-09-27 Amar Shah , Zoubin Ghahramani

Probabilistic generative models can be used for compression, denoising, inpainting, texture synthesis, semi-supervised learning, unsupervised feature learning, and other tasks. Given this wide range of applications, it is not surprising…

Machine Learning · Statistics 2016-04-26 Lucas Theis , Aäron van den Oord , Matthias Bethge
‹ Prev 1 3 4 5 6 7 10 Next ›