中文
相关论文

相关论文: On the distribution of cross-validated Mahalanobis…

200 篇论文

Mahalanobis distance is a classical tool in multivariate analysis. We suggest here an extension of this concept to the case of functional data. More precisely, the proposed definition concerns those statistical problems where the sample…

统计方法学 · 统计学 2018-03-20 José R. Berrendero , Beatriz Bueno-Larraz , Antonio Cuevas

Representational similarity analysis (RSA) tests models of brain computation by investigating how neural activity patterns reflect experimental conditions. Instead of predicting activity patterns directly, the models predict the geometry of…

This paper presents a general notion of Mahalanobis distance for functional data that extends the classical multivariate concept to situations where the observed data are points belonging to curves generated by a stochastic process. More…

统计理论 · 数学 2013-04-18 Esdras Joseph , Pedro Galeano , Rosa E. Lillo

The Mahalanobis distance is a classical tool used to measure the covariance-adjusted distance between points in $\bbR^d$. In this work, we extend the concept of Mahalanobis distance to separable Banach spaces by reinterpreting it as a…

机器学习 · 统计学 2025-10-31 Nikita Zozoulenko , Thomas Cass , Lukas Gonon

The concepts of similarity and distance are crucial in data mining. We consider the problem of defining the distance between two data sets by comparing summary statistics computed from the data sets. The initial definition of our distance…

数据结构与算法 · 计算机科学 2019-02-05 Nikolaj Tatti

A fundamental question in data analysis, machine learning and signal processing is how to compare between data points. The choice of the distance metric is specifically challenging for high-dimensional data sets, where the problem of…

机器学习 · 统计学 2017-08-15 Almog Lahav , Ronen Talmon , Yuval Kluger

Rerandomization, a design that utilizes pretreatment covariates and improves their balance between different treatment groups, has received attention recently in both theory and practice. From a survey by Bruhn and McKenzie (2009), there…

统计方法学 · 统计学 2025-09-17 Yuhao Wang , Xinran Li

Representational Similarity Analysis (RSA) is a popular method for analyzing neuroimaging and behavioral data. Here we evaluate the accuracy and reliability of RSA in the context of model selection, and compare it to that of regression.…

统计方法学 · 统计学 2025-11-18 Chuanji Gao , Gang Chen , Svetlana V. Shinkareva , Rutvik H. Desai

For many machine learning algorithms such as $k$-Nearest Neighbor ($k$-NN) classifiers and $ k $-means clustering, often their success heavily depends on the metric used to calculate distances between different data points. An effective…

计算机视觉与模式识别 · 计算机科学 2010-03-03 Chunhua Shen , Junae Kim , Lei Wang

Classical multivariate statistics measures the outlyingness of a point by its Mahalanobis distance from the mean, which is based on the mean and the covariance matrix of the data. A multivariate depth function is a function which, given a…

统计方法学 · 统计学 2021-05-06 Karl Mosler , Pavlo Mozharovskyi

Statistical analysis on non-Euclidean spaces typically relies on distances as the primary tool for constructing likelihoods. However, manifold-valued data admits richer structures in addition to Riemannian distances. We demonstrate that…

统计理论 · 数学 2026-03-25 Nicolas Escobar-Velasquez , Jaroslaw Harezlak

The classification of high dimensional data with kernel methods is considered in this article. Exploit- ing the emptiness property of high dimensional spaces, a kernel based on the Mahalanobis distance is proposed. The computation of the…

数值分析 · 计算机科学 2012-09-11 M. Fauvel , A. Villa , J. Chanussot , J. A. Benediktsson

Clustering and classification critically rely on distance metrics that provide meaningful comparisons between data points. We present mixed-integer optimization approaches to find optimal distance metrics that generalize the Mahalanobis…

机器学习 · 计算机科学 2018-03-29 Krishnan Kumaran , Dimitri Papageorgiou , Yutong Chang , Minhan Li , Martin Takáč

Multivariate equivalence testing is needed in a variety of scenarios for drug development. For example, drug products obtained from natural sources may contain many components for which the individual effects and/or their interactions on…

统计方法学 · 统计学 2024-06-07 Chao Wang , Yu-Ting Weng , Shaobo Liu , Tengfei Li , Meiyu Shen , Yi Tsong

A framework for assessing the matrix variate normality of three-way data is developed. The framework comprises a visual method and a goodness of fit test based on the Mahalanobis squared distance (MSD). The MSD of multivariate and matrix…

统计方法学 · 统计学 2019-10-08 Nikola Pocuca , Michael P. B. Gallaugher , Katharine M. Clark , Paul D. McNicholas

The need for appropriate ways to measure the distance or similarity between data is ubiquitous in machine learning, pattern recognition and data mining, but handcrafting such good metrics for specific problems is generally difficult. This…

机器学习 · 计算机科学 2019-01-25 Aurélien Bellet , Amaury Habrard , Marc Sebban

As regression is a widely studied problem, many methods have been proposed to solve it, each of them often requiring setting different hyper-parameters. Therefore, selecting the proper method for a given application may be very difficult…

机器学习 · 计算机科学 2026-03-23 Nassime Mountasir , Baptiste Lafabregue , Bruno Albert , Nicolas Lachiche

A collection of robust Mahalanobis distances for multivariate outlier detection is proposed, based on the notion of shrinkage. Robust intensity and scaling factors are optimally estimated to define the shrinkage. Some properties are…

统计方法学 · 统计学 2020-01-06 Elisa Cabana , Rosa E. Lillo , Henry Laniado

Metric learning makes it plausible to learn distances for complex distributions of data from labeled data. However, to date, most metric learning methods are based on a single Mahalanobis metric, which cannot handle heterogeneous data well.…

机器学习 · 统计学 2012-01-04 Caiming Xiong , David Johnson , Ran Xu , Jason J. Corso

Similarity metrics such as representational similarity analysis (RSA) and centered kernel alignment (CKA) have been used to compare layer-wise representations between neural networks. However, these metrics are confounded by the population…

机器学习 · 统计学 2022-02-02 Tianyu Cui , Yogesh Kumar , Pekka Marttinen , Samuel Kaski
‹ 上一页 1 2 3 10 下一页 ›