中文
相关论文

相关论文: Detecting British Columbia Coastal Rainfall Patter…

200 篇论文

Clustering is widely used in unsupervised learning to find homogeneous groups of observations within a dataset. However, clustering mixed-type data remains a challenge, as few existing approaches are suited for this task. This study…

机器学习 · 统计学 2025-11-26 Badih Ghattas , Alvaro Sanchez San-Benito

Industrial process monitoring increasingly relies on sensor-generated time-series data, yet the lack of labels, high variability, and operational noise make it difficult to extract meaningful patterns using conventional methods. Existing…

机器学习 · 计算机科学 2025-11-18 Zhipeng Ma , Bo Nørregaard Jørgensen , Zheng Grace Ma

Elastic Riemannian metrics have been used successfully in the past for statistical treatments of functional and curve shape data. However, this usage has suffered from an important restriction: the function boundaries are assumed fixed and…

统计方法学 · 统计学 2021-05-19 Darshan Bryner , Anuj Srivastava

The persistence over space and time of flash flood disasters -- flash floods that have caused either economical or life losses, or both -- is a diagnostic measure of areas subjected to hydrological risk. The concept of persistence can be…

应用统计 · 统计学 2021-07-28 Nan Wang , Luigi Lombardo , Marj Tonini , Weiming Cheng , Liang Guo , Junnan Xiong

In modeling spatial extremes, the dependence structure is classically inferred by assuming that block maxima derive from max-stable processes. Weather stations provide daily records rather than just block maxima. The point process approach…

统计方法学 · 统计学 2022-12-15 Hongwei Shang , Jun Yan , Xuebin Zhang

To facilitate effective decision-making, precipitation datasets should include uncertainty estimates. Quantile regression with machine learning has been proposed for issuing such estimates. Distributional regression offers distinct…

机器学习 · 计算机科学 2025-01-07 Georgia Papacharalampous , Hristos Tyralis , Nikolaos Doulamis , Anastasios Doulamis

A given set of data-points in some feature space may be associated with a Schrodinger equation whose potential is determined by the data. This is known to lead to good clustering solutions. Here we extend this approach into a full-fledged…

数据分析、统计与概率 · 物理学 2010-02-16 Marvin Weinstein , David Horn

A data filtering method for cluster analysis is proposed, based on minimizing a least squares function with a weighted $\ell_0$-norm penalty. To overcome the discontinuity of the objective function, smooth non-convex functions are employed…

最优化与控制 · 数学 2017-05-23 Andrea Cristofari

We present two methods for detecting patterns and clusters in high dimensional time-dependent functional data. Our methods are based on wavelet-based similarity measures, since wavelets are well suited for identifying highly discriminant…

统计方法学 · 统计学 2013-02-15 Anestis Antoniadis , Xavier Brossat , Jairo Cugliari , Jean-Michel Poggi

A set of probabilities along with corresponding quantiles are often used to define predictive distributions or probabilistic forecasts. These quantile predictions offer easily interpreted uncertainty of an event, and quantiles are generally…

统计方法学 · 统计学 2025-10-10 Spencer Wadsworth , Jarad Niemi

We propose a computationally efficient statistical method to obtain distributional properties of annual maximum 24 hour precipitation on a 1 km by 1 km regular grid over Iceland. A latent Gaussian model is built which takes into account…

统计方法学 · 统计学 2015-06-23 Óli Páll Geirsson , Birgir Hrafnkelsson , Daniel Simpson

We discuss the utility of applying clustering as a preprocessing step for identifying subseasonal to seasonal forecasts of opportunity of coastal sea level using convolutional neural networks (CNNs). Clustering leverages potential…

大气与海洋物理 · 物理学 2025-10-22 Laura Thapa , Marybeth Arcodia , Elizabeth A. Barnes

Clustering is one of the major tasks in data mining. In the last few years, Clustering of spatial data has received a lot of research attention. Spatial databases are components of many advanced information systems like geographic…

数据库 · 计算机科学 2012-06-04 Mohamed A. El-Zawawy

A task of clustering data given in the ordinal scale under conditions of overlapping clusters has been considered. It's proposed to use an approach based on memberhsip and likelihood functions sharing. A number of performed experiments…

机器学习 · 计算机科学 2017-02-07 Zhengbing Hu , Yevgeniy V. Bodyanskiy , Oleksii K. Tyshchenko , Viktoriia O. Samitova

Multiple imputation (MI) is a popular method for dealing with missing values. One main advantage of MI is to separate the imputation phase and the analysis one. However, both are related since they are based on distribution assumptions that…

统计方法学 · 统计学 2021-06-09 Vincent Audigier , Ndèye Niang , Matthieu Resche-Rigon

Streamflow, as a natural phenomenon, is continuous in time and so are the meteorological variables which influence its variability. In practice, it can be of interest to forecast the whole flow curve instead of points (daily or hourly). To…

应用统计 · 统计学 2016-10-20 Pierre Masselot , Sophie Dabo-Niang , Fateh Chebana , Taha B. M. J. Ouarda

Many common clustering methods cannot be used for clustering multivariate longitudinal data in cases where variables exhibit high autocorrelations. In this article, a copula kernel mixture model (CKMM) is proposed for clustering data of…

统计方法学 · 统计学 2025-06-23 Xi Zhang , Orla A. Murphy , Paul D. McNicholas

Meteorologists use shapes and movements of clouds in satellite images as indicators of several major types of severe storms. Satellite imaginary data are in increasingly higher resolution, both spatially and temporally, making it impossible…

计算机视觉与模式识别 · 计算机科学 2019-06-26 Xinye Zheng , Jianbo Ye , Yukun Chen , Stephen Wistar , Jia Li , Jose A. Piedra-Fernández , Michael A. Steinberg , James Z. Wang

The fundamental aim of clustering algorithms is to partition data points. We consider tasks where the discovered partition is allowed to vary with some covariate such as space or time. One approach would be to use fragmentation-coagulation…

机器学习 · 统计学 2013-11-01 Konstantina Palla , David A. Knowles , Zoubin Ghahramani

Common clustering algorithms require multiple scans of all the data to achieve convergence, and this is prohibitive when large databases, with data arriving in streams, must be processed. Some algorithms to extend the popular K-means method…

应用统计 · 统计学 2017-12-22 Giacomo Aletti , Alessandra Micheletti