English
Related papers

Related papers: Using an expert deviation carrying the knowledge o…

200 papers

Faithful uncertainty quantification (UQ) is paramount in high stakes climate prediction. Deep ensembles, or ensembles of probabilistic neural networks, are state of the art for UQ in machine learning (ML) and are growing increasingly…

Atmospheric and Oceanic Physics · Physics 2026-03-24 Devin M. McAfee , Elizabeth A. Barnes

We present a novel approach, in which we learn to cluster data directly from side information, in the form of a small set of pairwise examples. Unlike previous methods, with or without side information, we do not need to know the number of…

Machine Learning · Computer Science 2023-05-31 Michael A. Hobley , Victor A. Prisacariu

This report provides an exploration of different distance measures that can be used with the $K$-means algorithm for cluster analysis. Specifically, we investigate the Mahalanobis distance, and critically assess any benefits it may have…

Other Statistics · Statistics 2024-04-23 Zoe Shapcott

Purpose: The primary goal of this study is to explore the application of evaluation metrics to different clustering algorithms using the data provided from the Canadian Longitudinal Study (CLSA), focusing on cognitive features. The…

Machine Learning · Computer Science 2025-05-19 ChenNingZhi Sheng

We address the identification of grain-corresponding Laue reflections in energy dispersive Laue diffraction (EDLD) experiments by formulating it as a clustering problem solvable through unsupervised machine learning (ML). To achieve…

Through Ecological Momentary Assessment (EMA) studies, a number of time-series data is collected across multiple individuals, continuously monitoring various items of emotional behavior. Such complex data is commonly analyzed in an…

Machine Learning · Computer Science 2023-10-12 Mandani Ntekouli , Gerasimos Spanakis , Lourens Waldorp , Anne Roefs

Efficient Multimodal Large Language Models (MLLMs) compress vision tokens to reduce resource consumption, but the loss of visual information can degrade comprehension capabilities. Although some priors introduce Knowledge Distillation to…

Computer Vision and Pattern Recognition · Computer Science 2025-11-27 Ze Feng , Sen Yang , Boqiang Duan , Wankou Yang , Jingdong Wang

The widely applied k-means algorithm produces clusterings that violate our expectations with respect to high/low similarity/density and is in conflict with Kleinberg's axiomatic system for distance based clustering algorithms that…

Machine Learning · Computer Science 2023-08-08 Mieczysław A. Kłopotek

To quantify degree of spatial inhomogeneity for multiphase materials we adapt the entropic descriptor (ED) of a pillar model developed to greyscale images. To uncover the contribution of each phase we introduce the suitable 'phase…

Statistical Mechanics · Physics 2015-06-17 D. Fraczek , R. Piasecki

This paper explores hierarchical clustering in the case where pairs of points have dissimilarity scores (e.g. distances) as a part of the input. The recently introduced objective for points with dissimilarity scores results in every tree…

Machine Learning · Computer Science 2020-09-01 Benjamin Moseley , Yuyan Wang

4D-variational data assimilation is applied to the Lorenz '63 model to introduce a new method for parameter estimation in chaotic climate models. The approach aims to optimise an Earth system model (ESM), for which no adjoint exists, by…

Atmospheric and Oceanic Physics · Physics 2025-04-18 Philip David Kennedy , Abhirup Banerjee , Armin Köhl , Detlef Stammer

Recognizing subtle historical patterns is central to modeling and forecasting problems in time series analysis. Here we introduce and develop a new approach to quantify deviations in the underlying hidden generators of observed data…

Machine Learning · Statistics 2019-10-09 Yi Huang , Ishanu Chattopadhyay

Distance metric learning algorithms aim to appropriately measure similarities and distances between data points. In the context of clustering, metric learning is typically applied with the assist of side-information provided by experts,…

Machine Learning · Computer Science 2021-05-27 Rodrigo Randel , Daniel Aloise , Alain Hertz

Upcoming Sunyaev-Zel'dovich surveys are expected to return ~10^4 intermediate mass clusters at high redshift. Their average masses must be known to same accuracy as desired for the dark energy properties. Internal to the surveys, the CMB…

Astrophysics · Physics 2011-05-12 Wayne Hu , Simon DeDeo , Chris Vale

The impact of an extreme climate event depends strongly on its geographical scale. Max-stable processes can be used for the statistical investigation of climate extremes and their spatial dependencies on a continuous area. Most existing…

Methodology · Statistics 2023-06-14 Justus Contzen , Thorsten Dickhaus , Gerrit Lohmann

A novel methodology is proposed for clustering multivariate time series data using energy distance defined in Sz\'ekely and Rizzo (2013). Specifically, a dissimilarity matrix is formed using the energy distance statistic to measure…

Methodology · Statistics 2024-03-13 Richard A. Davis , Leon Fernandes , Konstantinos Fokianos

Unsupervised classification called clustering is a process of organizing objects into groups whose members are similar in some way. Clustering of uncertain data objects is a challenge in spatial data bases. In this paper we use Probability…

Databases · Computer Science 2013-12-10 Ramachandra Rao Kurada

K-means algorithm is a very popular clustering algorithm which is famous for its simplicity. Distance measure plays a very important rule on the performance of this algorithm. We have different distance measure techniques available. But…

Machine Learning · Computer Science 2014-05-30 Mr. Dibya Jyoti Bora , Dr. Anil Kumar Gupta

Time series clustering is an unsupervised learning method for classifying time series data into groups with similar behavior. It is used in applications such as healthcare, finance, economics, energy, and climate science. Several time…

Machine Learning · Statistics 2025-05-08 Chutiphan Charoensuk , Nathakhun Wiroonsri

Distance-based hierarchical clustering (HC) methods are widely used in unsupervised data analysis but few authors take account of uncertainty in the distance data. We incorporate a statistical model of the uncertainty through corruption or…

Machine Learning · Statistics 2016-09-02 Dekang Zhu , Dan P. Guralnik , Xuezhi Wang , Xiang Li , Bill Moran
‹ Prev 1 4 5 6 7 8 10 Next ›