English
Related papers

Related papers: Bayesian nonparametric modeling of dynamic polluti…

200 papers

Density-based clustering algorithms are widely used for discovering clusters in pattern recognition and machine learning since they can deal with non-hyperspherical clusters and are robustness to handle outliers. However, the runtime of…

Machine Learning · Computer Science 2022-07-07 Difei Cheng , Ruihang Xu , Bo Zhang , Ruinan Jin

We demonstrate that the clustering statistics and the corresponding phase transition to non-equilibrium clustering found in many experiments and simulation studies with self-propelled particles (SPPs) with alignment can be obtained from a…

Statistical Mechanics · Physics 2013-10-03 Fernando Peruani , Markus Baer

Fine particulate matter (PM2.5) is a mixture of air pollutants that has adverse effects on human health. Understanding the health effects of PM2.5 mixture and its individual species has been a research priority over the past two decades.…

Applications · Statistics 2019-09-10 Yawen Guan , Brian J Reich , James A Mulholland , Howard H Chang

Monitoring the atmospheric dispersion of pollutants is increasingly critical for environmental impact assessments. High-fidelity computational models are often employed to simulate plume dynamics, guiding decision-making and prioritizing…

Fluid Dynamics · Physics 2026-01-14 Ike Griss Salas , Megan R. Ebers , Jake Stevens-Haas , J. Nathan Kutz

Per- and polyfluoroalkyl substances (PFAS) are persistent environmental pollutants of major public health concern due to their resistance to degradation, widespread presence, and potential health risks. Analyzing PFAS in groundwater is…

Applications · Statistics 2026-03-17 Suman Majumder , Indranil Sahoo

We propose a general statistical framework for clustering multiple time series that exhibit nonlinear dynamics into an a-priori-unknown number of sub-groups. Our motivation comes from neuroscience, where an important problem is to identify,…

Machine Learning · Statistics 2019-03-05 Alexander Lin , Yingzhuo Zhang , Jeremy Heng , Stephen A. Allsop , Kay M. Tye , Pierre E. Jacob , Demba Ba

Discovering and clustering subspaces in high-dimensional data is a fundamental problem of machine learning with a wide range of applications in data mining, computer vision, and pattern recognition. Earlier methods divided the problem into…

Machine Learning · Statistics 2018-08-30 Maryam Jaberi , Marianna Pensky , Hassan Foroosh

In this article, an overview of Bayesian methods for sequential simulation from posterior distributions of nonlinear and non-Gaussian dynamic systems is presented. The focus is mainly laid on sequential Monte Carlo methods, which are based…

Methodology · Statistics 2023-04-28 Konstantinos E. Tatsis , Vasilis K. Dertimanis , Eleni N. Chatzi

Clustering is a crucial task in various domains of knowledge, including medicine, epidemiology, genomics, environmental science, economics, and visual sciences, among others. Methodologies for inferring the number of clusters have often…

Methodology · Statistics 2025-05-26 Clara Grazian

This paper studies a factor modeling-based approach for clustering high-dimensional data generated from a mixture of strongly correlated variables. Statistical modeling with correlated structures pervades modern applications in economics,…

Statistics Theory · Mathematics 2024-08-23 Shange Tang , Soham Jana , Jianqing Fan

Discrete mixture models are routinely used for density estimation and clustering. While conducting inferences on the cluster-specific parameters, current frequentist and Bayesian methods often encounter problems when clusters are placed too…

Methodology · Statistics 2012-09-21 Francesca Petralia , Vinayak Rao , David B. Dunson

A new model-based procedure is developed for sparse clustering of functional data that aims to classify a sample of curves into homogeneous groups while jointly detecting the most informative portions of domain. The proposed method is…

Methodology · Statistics 2023-10-04 Fabio Centofanti , Antonio Lepore , Biagio Palumbo

Statistical node clustering in discrete time dynamic networks is an emerging field that raises many challenges. Here, we explore statistical properties and frequentist inference in a model that combines a stochastic block model (SBM) for…

Methodology · Statistics 2016-06-23 Catherine Matias , Vincent Miele

Given a sample from a discretely observed compound Poisson process, we consider non-parametric estimation of the density $f_0$ of its jump sizes, as well as of its intensity $\lambda_0.$ We take a Bayesian approach to the problem and…

Statistics Theory · Mathematics 2023-02-27 Shota Gugushvili , Frank van der Meulen , Peter Spreij

Tracking and estimating Daily Fine Particulate Matter (PM2.5) is very important as it has been shown that PM2.5 is directly related to mortality related to lungs, cardiovascular system, and stroke. That is, high values of PM2.5 constitute a…

Methodology · Statistics 2019-09-06 Zhixing Xu , Jonathan R. Bradley , Debajyoti Sinha

Cluster analysis aims at partitioning data into groups or clusters. In applications, it is common to deal with problems where the number of clusters is unknown. Bayesian mixture models employed in such applications usually specify a…

Methodology · Statistics 2022-01-27 Jan Greve , Bettina Grün , Gertraud Malsiner-Walli , Sylvia Frühwirth-Schnatter

Genes are often regulated in living cells by proteins called transcription factors (TFs) that bind directly to short segments of DNA in close proximity to specific genes. These binding sites have a conserved nucleotide appearance, which is…

Statistics Theory · Mathematics 2007-06-13 Shane T. Jensen , Jun S. Liu

In Hermitian impurity scattering, each isolated late-time exponential is the fingerprint of a bound state. We show that this correspondence breaks down in non-Hermitian bands. For a single impurity in a non-Hermitian lattice, the late-time…

Mesoscale and Nanoscale Physics · Physics 2026-04-15 Ao Yang , Kai Zhang , Chen Fang

We present a novel approach to ecological risk assessment by recasting the Species Sensitivity Distribution (SSD) method within a Bayesian nonparametric (BNP) framework. Widely mandated by environmental regulatory bodies globally, SSD has…

Methodology · Statistics 2026-02-05 Louise Alamichel , Julyan Arbel , Guillaume Kon Kam King , Igor Prünster

Bayesian clustering accounts for uncertainty but is computationally demanding at scale. Furthermore, real-world datasets often contain missing values, and simple imputation ignores the associated uncertainty, resulting in suboptimal…

Machine Learning · Computer Science 2026-03-18 Prajit Bhaskaran , Tom Viering