English
Related papers

Related papers: Data mining and time series segmentation via extre…

200 papers

Time series classification is an important problem in data mining with several applications in different domains. Because time series data are usually high dimensional, dimensionality reduction techniques have been proposed as an efficient…

Machine Learning · Computer Science 2020-10-05 Muhammad Marwan Muhammad Fuad

In this paper, we propose a nonparametric approach that can be used in envelope extraction, peak-burst detection and clustering in time series. Our problem formalization results in a naturally defined splitting/forking of the time series.…

Machine Learning · Computer Science 2021-09-07 Kaan Gokcesu , Hakan Gokcesu

The characterisation of time-series data via their most salient features is extremely important in a range of machine learning task, not least of all with regards to classification and clustering. While there exist many feature extraction…

Machine Learning · Computer Science 2015-07-28 Duncan Barrack , James Goulding , Keith Hopcraft , Simon Preston , Gavin Smith

Data discretization, also known as binning, is a frequently used technique in computer science, statistics, and their applications to biological data analysis. We present a new method for the discretization of real-valued data into a finite…

Other Quantitative Biology · Quantitative Biology 2007-05-23 Elena S. Dimitrova , John J. McGee , Reinhard C. Laubenbacher

The extremal index $\theta$, a number in the interval $[0,1]$, is known to be a measure of primal importance for analyzing the extremes of a stationary time series. New rank-based estimators for $\theta$ are proposed which rely on the…

Statistics Theory · Mathematics 2020-06-30 Axel Bücher , Tobias Jennessen

In real world, the huge amount of temporal data is to be processed in many application areas such as scientific, financial, network monitoring, sensor data analysis. Data mining techniques are primarily oriented to handle discrete features.…

Databases · Computer Science 2014-02-19 P. Chaudhari , D. P. Rana , R. G. Mehta , N. J. Mistry , M. M. Raghuwanshi

We survey the application of a relatively new branch of statistical physics--"community detection"-- to data mining. In particular, we focus on the diagnosis of materials and automated image segmentation. Community detection describes the…

Materials Science · Physics 2017-11-22 Z. Nussinov , P. Ronhovde , Dandan Hu , S. Chakrabarty , M. Sahu , Bo Sun , N. A. Mauro , K. K. Sahu

We introduce Extrema-Segmented Entropy (ExSEnt), a feature-decomposed framework for quantifying time-series complexity that separates temporal from amplitude contributions. The method partitions a signal into monotonic segments by detecting…

Chaotic Dynamics · Physics 2025-09-30 Sara Kamali , Fabiano Baroni , Pablo Varona

Extreme environmental events frequently exhibit spatial and temporal dependence. These data are often modeled using max stable processes (MSPs). MSPs are computationally prohibitive to fit for as few as a dozen observations, with supposed…

Methodology · Statistics 2022-05-02 Emily C. Hector , Brian J. Reich

Discovering frequent episodes in event sequences is an interesting data mining task. In this paper, we argue that this framework is very effective for analyzing multi-neuronal spike train data. Analyzing spike train data is an important…

Databases · Computer Science 2008-03-10 Debprakash Patnaik , P. S. Sastry , K. P. Unnikrishnan

Compared to frequent pattern mining, sequential pattern mining emphasizes the temporal aspect and finds broad applications across various fields. However, numerous studies treat temporal events as single time points, neglecting their…

Databases · Computer Science 2025-07-18 Shuang Liang , Lili Chen , Wensheng Gan , Philip S. Yu , Shengjie Zhao

This paper proposes a novel kernel-based optimization scheme to handle tasks in the analysis, e.g., signal spectral estimation and single-channel source separation of 1D non-stationary oscillatory data. The key insight of our optimization…

Machine Learning · Statistics 2022-12-12 Jieren Xu , Yitong Li , Haizhao Yang , David Dunson , Ingrid Daubechies

Feature tracking is a common task in visualization applications, where methods based on topological data analysis (TDA) have successfully been applied in the past for feature definition as well as tracking. In this work, we focus on…

Graphics · Computer Science 2023-08-21 Emma Nilsson , Jonas Lukasczyk , Talha Bin Masood , Christoph Garth , Ingrid Hotz

There is an increasing need for algorithms that can accurately detect changepoints in long time-series, or equivalent, data. Many common approaches to detecting changepoints, for example based on penalised likelihood or minimum description…

Methodology · Statistics 2014-09-08 Robert Maidstone , Toby Hocking , Guillem Rigaill , Paul Fearnhead

This brief paper summarize the chances offered by the Peak-Over-Threshold method, related with analysis of extremes. Identification of appropriate Value at Risk can be solved by fitting data with a Generalized Pareto Distribution. Also an…

Applications · Statistics 2015-09-04 Gianluca Rosso

In this paper we deal with a network of agents seeking to solve in a distributed way Mixed-Integer Linear Programs (MILPs) with a coupling constraint (modeling a limited shared resource) and local constraints. MILPs are NP-hard problems and…

Systems and Control · Computer Science 2020-10-28 Andrea Camisa , Ivano Notarnicola , Giuseppe Notarstefano

We describe the $\texttt{Period detection and Identification Pipeline Suite}$ ($\texttt{PIPS}$) -- a new, fast, and statistically robust platform for period detection and analysis of astrophysical time-series data. $\texttt{PIPS}$ is an…

Binary segmentation is the classic greedy algorithm which recursively splits a sequential data set by optimizing some loss or likelihood function. Binary segmentation is widely used for changepoint detection in data sets measured over space…

Machine Learning · Computer Science 2024-10-14 Toby Dylan Hocking

Piecewise Aggregate Approximation (PAA) is a competitive basic dimension reduction method for high-dimensional time series mining. When deployed, however, the limitations are obvious that some important information will be missed,…

Machine Learning · Computer Science 2019-07-02 Chunkai Zhang , Yingyang Chen , Ao Yin , Zhen Qin , Xing Zhang , Keli Zhang , Zoe L. Jiang

We explore a combinatorial framework which efficiently quantifies the asymmetries between minima and maxima in local fluctuations of time series. We firstly showcase its performance by applying it to a battery of synthetic cases. We find…

Data Analysis, Statistics and Probability · Physics 2017-10-16 Uri Hasson , Jacopo Iacovacci , Ben Davis , Ryan Flanagan , Enzo Tagliazucchi , Helmut Laufs , Lucas Lacasa
‹ Prev 1 2 3 10 Next ›