中文
相关论文

相关论文: AWT -- Clustering Meteorological Time Series Using…

200 篇论文

Unsupervised Anomaly Detection (UAD) plays a crucial role in identifying abnormal patterns within data without labeled examples, holding significant practical implications across various domains. Although the individual contributions of…

机器学习 · 计算机科学 2024-06-04 Zeyu Fang , Ming Gu , Sheng Zhou , Jiawei Chen , Qiaoyu Tan , Haishuai Wang , Jiajun Bu

How does soil pollution affect a plant's circadian clock? Are there any differences between how the clock reacts when exposed to different concentrations of elements of the periodic table? If so, can we characterise these differences? We…

应用统计 · 统计学 2016-08-01 Jessica K. Hargreaves , Marina I. Knight , Jon W. Pitchford , Seth J. Davis

Improving the future of healthcare starts by better understanding the current actual practices in hospital settings. This motivates the objective of discovering typical care pathways from patient data. Revealing typical care pathways can be…

机器学习 · 计算机科学 2024-12-20 Thomas Guyet , Pierre Pinson , Enoal Gesny

Functional data present unique challenges for clustering due to their infinite-dimensional nature and potential sensitivity to outliers. An extension of the OCLUST algorithm to the functional setting is proposed to address these issues. The…

机器学习 · 统计学 2025-08-06 Katharine M. Clark , Paul D. McNicholas

Clustering has many important applications in computer science, but real-world datasets often contain outliers. Moreover, the presence of outliers can make the clustering problems to be much more challenging. To reduce the complexities,…

数据结构与算法 · 计算机科学 2020-05-04 Hu Ding , Jiawei Huang , Haikuo Yu

Clustering provides a common means of identifying structure in complex data, and there is renewed interest in clustering as a tool for the analysis of large data sets in many fields. A natural question is how many clusters are appropriate…

数据分析、统计与概率 · 物理学 2007-05-23 Susanne Still , William Bialek

We present two methods for detecting patterns and clusters in high dimensional time-dependent functional data. Our methods are based on wavelet-based similarity measures, since wavelets are well suited for identifying highly discriminant…

统计方法学 · 统计学 2013-02-15 Anestis Antoniadis , Xavier Brossat , Jairo Cugliari , Jean-Michel Poggi

Multi-view clustering has gained broad attention owing to its capacity to exploit complementary information across multiple data views. Although existing methods demonstrate delightful clustering performance, most of them are of high time…

机器学习 · 计算机科学 2023-03-06 Xinhang Wan , Xinwang Liu , Jiyuan Liu , Siwei Wang , Yi Wen , Weixuan Liang , En Zhu , Zhe Liu , Lu Zhou

Medical data often exhibit characteristics that make cluster analysis particularly challenging, such as missing values, outliers, and cluster features like skewness. Typically, such data would need to be preprocessed -- by cleaning outliers…

统计方法学 · 统计学 2025-12-16 Jason Pillay , Cristina Tortora , Antonio Punzo , Andriette Bekker

The k-means clustering algorithm is a popular algorithm that partitions data into k clusters. There are many improvements to accelerate the standard algorithm. Most current research employs upper and lower bounds on point-to-cluster…

机器学习 · 计算机科学 2024-10-22 Andreas Lang , Erich Schubert

A natural way to characterize the cluster structure of a dataset is by finding regions containing a high density of data. This can be done in a nonparametric way with a kernel density estimate, whose modes and hence clusters can be found…

机器学习 · 计算机科学 2015-03-03 Miguel Á. Carreira-Perpiñán

Ad-hoc networks are specifically designed to facilitate communication in environments where establishing a dedicated network infrastructure is exceedingly complex or impractical. The integration of clustering concepts into various ad-hoc…

网络与互联网体系结构 · 计算机科学 2024-07-04 Adda Boualem , Marwane Ayaida , Hichem Sedjelmaci , Chaimaa Khalfi , Kamilia Brahimi , Bochra Khelil , Sanaa Bouchama

We present a new framework to detect various types of variable objects within massive astronomical time-series data. Assuming that the dominant population of objects is non-variable, we find outliers from this population by using a…

天体物理仪器与方法 · 物理学 2010-01-17 Min-Su Shin , Michael Sekora , Yong-Ik Byun

Many automated systems need the capability of automatic change detection without the given detection threshold. This paper presents an automated change detection algorithm in streaming multivariate data. Two overlapping windows are used to…

数据库 · 计算机科学 2013-11-05 Dang-Hoan Tran

In tropical countries with high humidity, air conditioning can account for up to 60% of a building's energy use. For commercial buildings with centralized systems, the efficiency of the chiller plant is vital, and model predictive control…

系统与控制 · 电气工程与系统科学 2025-02-25 Zhan Wang , Chen Weidong , Huang Zhifeng , Md Raisul Islam , Chua Kian Jon

Hydroclimatic time series analysis focuses on a few feature types (e.g., autocorrelations, trends, extremes), which describe a small portion of the entire information content of the observations. Aiming to exploit a larger part of the…

We propose a new clustering technique that can be regarded as a numerical method to compute the proximity gestalt. The method analyzes edge length statistics in the MST of the dataset and provides an a contrario cluster detection criterion.…

机器学习 · 计算机科学 2011-07-20 Mariano Tepper , Pablo Musé , Andrés Almansa

One emerging application of machine learning methods is the inference of galaxy cluster masses. In this note, machine learning is used to directly combine five simulated multiwavelength measurements in order to find cluster masses. This is…

宇宙学与河外天体物理 · 物理学 2020-01-08 J. D. Cohn , Nicholas Battaglia

Clustering algorithms have long been the topic of research, representing the more popular side of unsupervised learning. Since clustering analysis is one of the best ways to find some clarity and structure within raw data, this paper…

机器学习 · 计算机科学 2025-11-25 Naitik Gada

Clustering algorithms are one of the main analytical methods to detect patterns in unlabeled data. Existing clustering methods typically treat samples in a dataset as points in a metric space and compute distances to group together similar…

机器学习 · 计算机科学 2021-10-12 Tarek Naous , Srinjay Sarkar , Abubakar Abid , James Zou