中文
相关论文

相关论文: Asymptotic analysis of the Gaussian kernel matrix …

200 篇论文

This paper discusses a special kind of a simple yet possibly powerful algorithm, called single-kernel Gradraker (SKG), which is an adaptive learning method predicting unknown nodal values in a network using known nodal values and the…

信号处理 · 电气工程与系统科学 2022-04-28 Yue Zhao , Ender Ayanoglu

In this paper, we consider the problem of testing equality of the covariance matrices of L complex Gaussian multivariate time series of dimension $M$ . We study the special case where each of the L covariance matrices is modeled as a rank K…

统计理论 · 数学 2024-04-11 Rémi Beisson , Pascal Vallet , Audrey Giremus , Guillaume Ginolhac

In this article we perform an asymptotic analysis of Bayesian parallel kernel density estimators introduced by Neiswanger, Wang and Xing (2014). We derive the asymptotic expansion of the mean integrated squared error for the full data…

统计理论 · 数学 2020-11-09 Alexey Miroshnikov , Evgeny Savelev

We propose a novel kernel-based nonparametric two-sample test, employing the combined use of kernel mean and kernel covariance embedding. Our test builds on recent results showing how such combined embeddings map distinct probability…

机器学习 · 统计学 2025-09-16 Leonardo V. Santoro , Victor M. Panaretos

Approximating significance scans of searches for new particles in high-energy physics experiments as Gaussian fields is a well-established way to estimate the trials factors required to quantify global significances. We propose a novel,…

数据分析、统计与概率 · 物理学 2023-10-23 V. Ananiev , A. L. Read

This paper considers a noisy data structure recovery problem. The goal is to investigate the following question: Given a noisy observation of a permuted data set, according to which permutation was the original data sorted? The focus is on…

信息论 · 计算机科学 2020-11-24 Minoh Jeong , Alex Dytso , Martina Cardone , H. Vincent Poor

Reliable prediction of protein variant effects is crucial for both protein optimization and for advancing biological understanding. For practical use in protein engineering, it is important that we can also provide reliable uncertainty…

生物大分子 · 定量生物学 2024-11-01 Peter Mørch Groth , Mads Herbert Kerrn , Lars Olsen , Jesper Salomon , Wouter Boomsma

This paper studies the optimality of kernel methods in high-dimensional data clustering. Recent works have studied the large sample performance of kernel clustering in the high-dimensional regime, where Euclidean distance becomes less…

机器学习 · 统计学 2019-12-03 Leena Chennuru Vankadara , Debarghya Ghoshdastidar

Pseudospectral analysis is fundamental for quantifying the sensitivity and transient behavior of nonnormal matrices, yet its computational cost scales cubically with dimension, rendering it prohibitive for large-scale systems. While…

数值分析 · 数学 2026-02-03 Vladimir R. Kostic , Dragana Lj. Cvetkovic , Ljiljana Cvetkovic

Modern large scale datasets are often plagued with missing entries. For tabular data with missing values, a flurry of imputation algorithms solve for a complete matrix which minimizes some penalized reconstruction error. However, almost…

机器学习 · 统计学 2021-01-20 Yuxuan Zhao , Madeleine Udell

We introduce new Gaussian Process (GP) high-order approximations to linear operations that are frequently used in various numerical methods. Our method employs the kernel-based GP regression modeling, a non-parametric Bayesian approach to…

计算物理 · 物理学 2025-06-09 Christopher DeGrendele , Dongwook Lee

We consider ensembles of Gaussian (Hermite) and Wishart (Laguerre) $N\times N$ hermitian matrices. We study the effect of finite rank perturbations of these ensembles by a source term. The rank $r$ of the perturbation corresponds to the…

数学物理 · 物理学 2007-05-23 Patrick Desrosiers , Peter J. Forrester

Gaussian Process (GP) regression is a powerful nonparametric Bayesian framework, but its performance depends critically on the choice of covariance kernel. Selecting an appropriate kernel is therefore central to model quality, yet remains…

机器学习 · 计算机科学 2026-01-14 Md Shafiqul Islam , Shakti Prasad Padhy , Douglas Allaire , Raymundo Arróyave

A common feature of high-dimensional data is that the data dimension is high, however, the sample size is relatively low. We call such data HDLSS data. In this paper, we study asymptotic properties of the first principal component in the…

统计理论 · 数学 2015-03-26 Aki Ishii , Kazuyoshi Yata , Makoto Aoshima

Gaussian processes (GP) are Bayesian non-parametric models that are widely used for probabilistic regression. Unfortunately, it cannot scale well with large data nor perform real-time predictions due to its cubic time cost in the data size.…

机器学习 · 计算机科学 2014-08-12 Jie Chen , Nannan Cao , Kian Hsiang Low , Ruofei Ouyang , Colin Keng-Yan Tan , Patrick Jaillet

Gaussian processes (GP) are Bayesian non-parametric models that are widely used for probabilistic regression. Unfortunately, it cannot scale well with large data nor perform real-time predictions due to its cubic time cost in the data size.…

机器学习 · 统计学 2013-05-27 Jie Chen , Nannan Cao , Kian Hsiang Low , Ruofei Ouyang , Colin Keng-Yan Tan , Patrick Jaillet

Fisher's linear discriminant analysis is a classical method for classification, yet it is limited to capturing linear features only. Kernel discriminant analysis as an extension is known to successfully alleviate the limitation through a…

机器学习 · 统计学 2022-07-29 Jiae Kim , Yoonkyung Lee , Zhiyu Liang

Despite the increasing importance of stochastic processes on linear networks and graphs, current literature on multivariate (vector-valued) Gaussian random fields on metric graphs is elusive. This paper challenges several aspects related to…

统计理论 · 数学 2025-01-20 Tobia Filosi , Emilio Porcu , Xavier Emery , Claudio Agostinelli , Alfredo Alegrìa

Topological data analysis (TDA) is an emerging mathematical concept for characterizing shapes in complex data. In TDA, persistence diagrams are widely recognized as a useful descriptor of data, and can distinguish robust and noisy…

代数拓扑 · 数学 2016-04-27 Genki Kusano , Kenji Fukumizu , Yasuaki Hiraoka

Learning the kernel parameters for Gaussian processes is often the computational bottleneck in applications such as online learning, Bayesian optimization, or active learning. Amortizing parameter inference over different datasets is a…

机器学习 · 计算机科学 2023-06-19 Matthias Bitzer , Mona Meister , Christoph Zimmer