中文
相关论文

相关论文: Clustering multivariate functional data using the …

200 篇论文

Factor analysis has been extensively used to reveal the dependence structures among multivariate variables, offering valuable insight in various fields. However, it cannot incorporate the spatial heterogeneity that is typically present in…

统计方法学 · 统计学 2024-11-14 Yanxiu Jin , Tomoya Wakayama , Renhe Jiang , Shonosuke Sugasawa

Climate models lack the necessary resolution for urban climate studies, requiring computationally intensive processes to estimate high resolution air temperatures. In contrast, Data-driven approaches offer faster and more accurate air…

大气与海洋物理 · 物理学 2024-09-05 Fatemeh Chajaei , Hossein Bagheri

We propose a method called integrated diffusion for combining multimodal datasets, or data gathered via several different measurements on the same system, to create a joint data diffusion operator. As real world data suffers from both local…

机器学习 · 计算机科学 2022-03-07 Manik Kuchroo , Abhinav Godavarthi , Alexander Tong , Guy Wolf , Smita Krishnaswamy

We propose new ensemble models for multivariate functional data classification as combinations of semi-metric-based weak learners. Our models extend current semi-metric-type methods from the univariate to the multivariate case, propose new…

Clinical time series data are critical for patient monitoring and predictive modeling. These time series are typically multivariate and often comprise hundreds of heterogeneous features from different data sources. The grouping of features…

机器学习 · 计算机科学 2025-11-12 Fedor Sergeev , Manuel Burger , Polina Leshetkina , Vincent Fortuin , Gunnar Rätsch , Rita Kuznetsova

Air traffic controllers benefit from referencing historical dates with similar complex air traffic conditions to identify potential management measures and their effects, which is critical for understanding air transportation system laws…

应用统计 · 统计学 2025-09-12 Wei Sun , Zi-Feng Yi , Zhi-Qiang Feng , Ji Ma , Ruo-shi Yang

This paper introduces {\em fusion subspace clustering}, a novel method to learn low-dimensional structures that approximate large scale yet highly incomplete data. The main idea is to assign each datum to a subspace of its own, and minimize…

机器学习 · 计算机科学 2022-05-24 Usman Mahmood , Daniel Pimentel-Alarcón

Cluster analysis across multiple institutions poses significant challenges due to data-sharing restrictions. To overcome these limitations, we introduce the Federated One-shot Ensemble Clustering (FONT) algorithm, a novel solution tailored…

机器学习 · 统计学 2024-09-16 Rui Duan , Xin Xiong , Jueyi Liu , Katherine P. Liao , Tianxi Cai

Clustering is a well-established technique in machine learning and data analysis, widely used across various domains. Cluster validity indices, such as the Average Silhouette Width, Calinski-Harabasz, and Davies-Bouldin indices, play a…

机器学习 · 计算机科学 2026-04-16 Renato Cordeiro de Amorim , Vladimir Makarenkov

A recent study by Panchagnula et al. [J. Chem. Phys. 161, 054308 (2024)] illustrated the non-concordance of a variety of electronic structure methods at describing the symmetric double-well potential expected along the anisotropic direction…

化学物理 · 物理学 2025-10-28 K. Panchagnula , D. Graf , K. R. Bryenton , D. P. Tew , E. R. Johnson , A. J. W. Thom

The generation of initial conditions via accurate data assimilation is crucial for weather forecasting and climate modeling. We propose DiffDA as a denoising diffusion model capable of assimilating atmospheric variables using predicted…

计算工程、金融与科学 · 计算机科学 2024-06-11 Langwen Huang , Lukas Gianinazzi , Yuejiang Yu , Peter D. Dueben , Torsten Hoefler

We introduce a mathematical formulation of feature-informed data assimilation (FIDA). In FIDA, the information about feature events, such as shock waves, level curves, wavefronts and peak value, in dynamical systems are used for the…

系统与控制 · 电气工程与系统科学 2022-11-02 Wei Kang , Daniel M. Tartakovsky , Apoorv Srivastava

Functional data analysis is a statistical framework where data are assumed to follow some functional form. This method of analysis is commonly applied to time series data, where time, measured continuously or in discrete intervals, serves…

应用统计 · 统计学 2020-04-07 Forrest Paton , Paul D. McNicholas

Cosmic demographics -- the statistical study of populations of astrophysical objects -- has long relied on *multivariate statistics*, providing methods for analyzing data comprising fixed-length vectors of properties of objects, as might be…

天体物理仪器与方法 · 物理学 2024-08-27 Thomas Loredo , Tamas Budavari , David Kent , David Ruppert

A new method for clustering functional data is proposed via information maximization. The proposed method learns a probabilistic classifier in an unsupervised manner so that mutual information (or squared loss mutual information) between…

应用统计 · 统计学 2023-06-08 Xinyu Li , Jianjun Xu , Haoyang Cheng

A new model-based procedure is developed for sparse clustering of functional data that aims to classify a sample of curves into homogeneous groups while jointly detecting the most informative portions of domain. The proposed method is…

统计方法学 · 统计学 2023-10-04 Fabio Centofanti , Antonio Lepore , Biagio Palumbo

Diet is a risk factor for many diseases. In nutritional epidemiology, studying reproducible dietary patterns is critical to reveal important associations with health. However, it is challenging: diverse cultural and ethnic backgrounds may…

应用统计 · 统计学 2025-02-10 Roberta De Vito , Alejandra Avalos-Pacheco

Clustering algorithms partition a dataset into groups of similar points. The clustering problem is very general, and different partitions of the same dataset could be considered correct and useful. To fully understand such data, it must be…

机器学习 · 计算机科学 2021-02-02 James M. Murphy , Sam L. Polk

Multi-dimensional functional data arises in numerous modern scientific experimental and observational studies. In this paper we focus on longitudinal functional data, a structured form of multidimensional functional data. Operating within a…

统计方法学 · 统计学 2019-09-20 John Shamshoian , Damla Senturk , Shafali Jeste , Donatello Telesca

Clustering is essential in data analysis and machine learning, but traditional algorithms like $k$-means and Gaussian Mixture Models (GMM) often fail with nonconvex clusters. To address the challenge, we introduce the Flexible Bivariate…

机器学习 · 计算机科学 2025-02-28 Yung-Peng Hsu , Hung-Hsuan Chen