English
Related papers

Related papers: A Multivariate Extreme Value Theory Approach to An…

200 papers

A model based clustering procedure for data of mixed type, clustMD, is developed using a latent variable model. It is proposed that a latent variable, following a mixture of Gaussian distributions, generates the observed data of mixed type.…

Methodology · Statistics 2015-11-06 Damien McParland , Isobel Claire Gormley

In recent years, there has been a growing interest in identifying anomalous structure within multivariate data streams. We consider the problem of detecting collective anomalies, corresponding to intervals where one or more of the data…

Methodology · Statistics 2019-09-05 Alexander T M Fisch , Idris A Eckley , Paul Fearnhead

Healthcare cost prediction is a challenging task due to the high-dimensionality and high correlation among covariates. Additionally, the skewed, heavy-tailed, and often multi-modal nature of cost data can complicate matters further due to…

Methodology · Statistics 2023-03-13 Zhengxiao Li , Yifan Huang , Yang Cao

We extend the scope of the dynamical theory of extreme values to cover phenomena that do not happen instantaneously, but evolve over a finite, albeit unknown at the onset, time interval. We consider complex dynamical systems, composed of…

Neurons and Cognition · Quantitative Biology 2020-05-20 Theophile Caby , Giorgio Mantica

One of the main goal of extreme value analysis is to estimate the probability of rare events given a sample from an unknown distribution. The upper tail behavior of this distribution is described by the extreme value index. We present a new…

Probability · Mathematics 2007-05-23 Laurent Gardes , Stephane Girard

The $k$-means clustering algorithm and its variant, the spherical $k$-means clustering, are among the most important and popular methods in unsupervised learning and pattern detection. In this paper, we explore how the spherical $k$-means…

Methodology · Statistics 2019-05-28 Anja Janßen , Phyllis Wan

The conventional use of the Generalized Extreme Value (GEV) distribution to model block maxima may be inappropriate when extremes are actually structured into multiple heterogeneous groups. In this work, we propose a novel approach for…

Unsupervised Anomaly Detection (UAD) plays a crucial role in identifying abnormal patterns within data without labeled examples, holding significant practical implications across various domains. Although the individual contributions of…

Machine Learning · Computer Science 2024-06-04 Zeyu Fang , Ming Gu , Sheng Zhou , Jiawei Chen , Qiaoyu Tan , Haishuai Wang , Jiajun Bu

To capture the extremal behaviour of complex environmental phenomena in practice, flexi\-ble techniques for modelling tail behaviour are required. In this paper, we introduce a variety of such methods, which were used by the Lancopula…

We use extreme value theory to estimate the probability of successive exceedances of a threshold value of a time-series of an observable on several classes of chaotic dynamical systems. The observables have either a Fr\'echet (fat-tailed)…

Dynamical Systems · Mathematics 2023-11-07 Meagan Carney , Mark Holland , Matthew Nicol , Phuong Tran

In this paper, a scale mixture of Normal distributions model is developed for classification and clustering of data having outliers and missing values. The classification method, based on a mixture model, focuses on the introduction of…

Machine Learning · Statistics 2017-11-23 G. Revillon , A. Djafari , C. Enderli

We describe the applications of clustering and visualization tools using the so-called neutral B anomalies as an example. Clustering permits parameter space partitioning into regions that can be separated with some given measurements. It…

Data Analysis, Statistics and Probability · Physics 2023-04-04 Ursula Laa , German Valencia

We re-consider Leadbetter's extremal index for stationary sequences. It has interpretation as reciprocal of the expected size of an extremal cluster above high thresholds. We focus on heavy-tailed time series, in particular on regularly…

Probability · Mathematics 2021-06-10 Gloria Buriticá , Meyer Nicolas , Thomas Mikosch , Olivier Wintenberger

The use of video-imaging data for in-line process monitoring applications has become more and more popular in the industry. In this framework, spatio-temporal statistical process monitoring methods are needed to capture the relevant…

Applications · Statistics 2020-04-24 Hao Yan , Marco Grasso , Kamran Paynabar , Bianca Maria Colosimo

Combining strengths from deep learning and extreme value theory can help describe complex relationships between variables where extreme events have significant impacts (e.g., environmental or financial applications). Neural networks learn…

Applications · Statistics 2023-10-06 Mitchell L. Krock , Julie Bessac , Michael L. Stein

Modeling univariate block maxima by the generalized extreme value distribution constitutes one of the most widely applied approaches in extreme value statistics. It has recently been found that, for an underlying stationary time series,…

Statistics Theory · Mathematics 2021-11-01 Axel Bücher , Leandra Zanger

This paper unifies and extends results on a class of multivariate Extreme Value (EV) models studied by Hougaard, Crowder, and Tawn. In these models both unconditional and conditional distributions are EV, and all lower-dimensional marginals…

Methodology · Statistics 2013-09-30 Anne-Laure Fougères , John P. Nolan , Holger Rootzén

The conditional extremes (CE) framework has proven useful for analysing the joint tail behaviour of random vectors. However, when applied across many locations or variables, it can be difficult to interpret or compare the resulting extremal…

Methodology · Statistics 2025-10-24 Patrick O'Toole , Christian Rohrbeck , Jordan Richards

Mixed membership models are an extension of finite mixture models, where each observation can partially belong to more than one mixture component. A probabilistic framework for mixed membership models of high-dimensional continuous data is…

Learning how to rank multivariate unlabeled observations depending on their degree of abnormality/novelty is a crucial problem in a wide range of applications. In practice, it generally consists in building a real valued "scoring" function…

Machine Learning · Statistics 2015-02-06 Nicolas Goix , Anne Sabourin , Stéphan Clémençon