English
Related papers

Related papers: Estimation of cluster functionals for regularly va…

200 papers

Clustered standard errors and approximate randomization tests are popular inference methods that allow for dependence within observations. However, they require researchers to know the cluster structure ex ante. We propose a procedure to…

Econometrics · Economics 2022-01-14 Yong Cai

A central limit theorem is proved for some strictly stationary sequences of random variables that satisfy certain mixing conditions and are subjected to the "shrinking operators" $U_r(x):=[\max\{|x|-r,0\}]\cdot x/|x|,\ r \ge 0$. For…

Probability · Mathematics 2014-10-02 Richard C. Bradley , Zbigniew J. Jurek

Clustering algorithms frequently require the number of clusters to be chosen in advance, but it is usually not clear how to do this. To tackle this challenge when clustering within sequential data, we present a method for estimating the…

Machine Learning · Statistics 2024-07-29 Thomas van Vuren , Thomas Cronk , Jaron Sanders

High-order clustering aims to identify heterogeneous substructures in multiway datasets that arise commonly in neuroimaging, genomics, social network studies, etc. The non-convex and discontinuous nature of this problem pose significant…

Methodology · Statistics 2022-10-11 Rungang Han , Yuetian Luo , Miaoyan Wang , Anru R. Zhang

The random cluster model is used to define an upper bound on a distance measure as a function of the number of data points to be classified and the expected value of the number of classes to form in a hybrid K-means and regression…

Machine Learning · Computer Science 2016-02-12 Robert A. Murphy

We propose a mixture of latent trait models with common slope parameters (MCLT) for model-based clustering of high-dimensional binary data, a data type for which few established methods exist. Recent work on clustering of binary data, based…

Methodology · Statistics 2017-10-09 Yang Tang , Ryan P. Browne , Paul D. McNicholas

This article develops design-based ratio estimators for clustered, blocked randomized controlled trials (RCTs), with an application to a federally funded, school-based RCT testing the effects of behavioral health interventions. We consider…

Methodology · Statistics 2024-05-31 Peter Z. Schochet , Nicole E. Pashley , Luke W. Miratrix , Tim Kautz

Clustering coefficient is one of the most useful indices in complex networks. However, graph theoretic properties of this metric have not been discussed much in the literature, especially in graphs resulting from some binary operations. In…

Combinatorics · Mathematics 2022-04-20 Remarl Joseph M. Damalerio , Rolito G. Eballe

We develop a new method to find the number of volatility regimes in a nonstationary financial time series by applying unsupervised learning to its volatility structure. We use change point detection to partition a time series into locally…

Statistical Finance · Quantitative Finance 2022-11-15 Arjun Prakash , Nick James , Max Menzies , Gilad Francis

Clustering the nodes of a graph allows the analysis of the topology of a network. The stochastic block model is a clustering method based on a probabilistic model. Initially developed for binary networks it has recently been extended to…

Computation · Statistics 2014-02-17 Jean-Benoist Leger

We study hierarchical clusterings of metric spaces that change over time. This is a natural geometric primitive for the analysis of dynamic data sets. Specifically, we introduce and study the problem of finding a temporally coherent…

Data Structures and Algorithms · Computer Science 2017-10-23 Tamal K. Dey , Alfred Rossi , Anastasios Sidiropoulos

Clustering uncertain data has emerged as a challenging task in uncertain data management and mining. Thanks to a computational complexity advantage over other clustering paradigms, partitional clustering has been particularly studied and a…

Databases · Computer Science 2012-03-30 Francesco Gullo , Andrea Tagarelli

Cluster analysis is used to explore structure in unlabeled data sets in a wide range of applications. An important part of cluster analysis is validating the quality of computationally obtained clusters. A large number of different internal…

Machine Learning · Statistics 2018-01-10 Masud Moshtaghi , James C. Bezdek , Sarah M. Erfani , Christopher Leckie , James Bailey

This paper focuses on vector-valued composite functionals, which may be nonlinear in probability. Our primary goal is to establish central limit theorems for these functionals when mixed estimators are employed. Our study is relevant to the…

Statistics Theory · Mathematics 2025-01-09 Huihui Chen , Darinka Dentcheva , Yang Lin , Gregory J. Stock

Recent years have seen an increased focus into the tasks of predicting hospital inpatient risk of deterioration and trajectory evolution due to the availability of electronic patient data. A common approach to these problems involves…

Machine Learning · Computer Science 2020-11-18 Henrique Aguiar , Mauro Santos , Peter Watkinson , Tingting Zhu

Dynamical scaling and ageing in disordered systems far from equilibrium is reviewed. Particular attention is devoted to the question to what extent a recently introduced generalization of dynamical scaling to local scale-invariance can…

Statistical Mechanics · Physics 2007-05-23 Malte Henkel , Michel Pleimling

Spreading processes on graphs arise in a host of application domains, from the study of online social networks to viral marketing to epidemiology. Various discrete-time probabilistic models for spreading processes have been proposed. These…

Social and Information Networks · Computer Science 2021-09-24 Abram Magner , Carolyn Kaminski , Petko Bogdanov

In the analysis of binary longitudinal data, it is of interest to model a dynamic relationship between a response and covariates as a function of time, while also investigating similar patterns of time-dependent interactions. We present a…

Methodology · Statistics 2023-04-11 Jinwon Sohn , Seonghyun Jeong , Young Min Cho , Taeyoung Park

The goal of lifetime clustering is to develop an inductive model that maps subjects into $K$ clusters according to their underlying (unobserved) lifetime distribution. We introduce a neural-network based lifetime clustering model that can…

Machine Learning · Computer Science 2019-10-03 S Chandra Mouli , Leonardo Teixeira , Jennifer Neville , Bruno Ribeiro

In the present paper, we studied a Dynamic Stochastic Block Model (DSBM) under the assumptions that the connection probabilities, as functions of time, are smooth and that at most $s$ nodes can switch their class memberships between two…

Methodology · Statistics 2017-05-04 Marianna Pensky , Teng Zhang