English
Related papers

Related papers: A spatial scan statistic for zero-inflated Poisson…

200 papers

Many recent studies have demonstrated that scaling arguments, such as the so-called hierarchical {\em ansatz}, are extremely useful in understanding the statistical properties of weak gravitational lensing. This is especially true on small…

Astrophysics · Physics 2009-10-31 Dipak Munshi , Peter Coles

Subspace clustering refers to the problem of segmenting high dimensional data drawn from a union of subspaces into the respective subspaces. In some applications, partial side-information to indicate "must-link" or "cannot-link" in…

Computer Vision and Pattern Recognition · Computer Science 2018-05-23 Chun-Guang Li , Junjian Zhang , Jun Guo

We introduce a novel statistical significance-based approach for clustering hierarchical data using semi-parametric linear mixed-effects models designed for responses with laws in the exponential family (e.g., Poisson and Bernoulli). Within…

Methodology · Statistics 2025-02-04 Alessandra Ragni , Chiara Masci , Francesca Ieva , Anna Maria Paganoni

A general approach to selective inference is considered for hypothesis testing of the null hypothesis represented as an arbitrary shaped region in the parameter space of multivariate normal model. This approach is useful for hierarchical…

Statistics Theory · Mathematics 2018-03-28 Yoshikazu Terada , Hidetoshi Shimodaira

The measurement of the abundance of galaxy clusters in the Universe is a sensitive probe of cosmology, which depends on both the expansion history of the Universe and the growth of structure. Density fluctuations across the finite survey…

Cosmology and Nongalactic Astrophysics · Physics 2024-06-19 Constantin Payerne , Calum Murray , Céline Combet , Mariana Penna-Lima

Compressed sensing is a technique for recovering an unknown sparse signal from a small number of linear measurements. When the measurement matrix is random, the number of measurements required for perfect recovery exhibits a phase…

Optimization and Control · Mathematics 2016-12-30 Mateo Díaz , Mauricio Junca , Felipe Rincón , Mauricio Velasco

In this paper, we propose a Bayesian Graphical LASSO for correlated countable data and apply it to spatial crime data. In the proposed model, we assume a Gaussian Graphical Model for the latent variables which dominate the potential risks…

Methodology · Statistics 2020-06-08 Sho Ichigozaki , Takahiro Kawashima , Hayaru Shouno

Count data occur widely in many bio-surveillance and healthcare applications, e.g., the numbers of new patients of different types of infectious diseases from different cities/counties/states repeatedly over time, say, daily/weekly/monthly.…

Applications · Statistics 2022-10-11 Yujie Zhao , Xiaoming Huo , Yajun Mei

We introduce the Poisson tensor completion (PTC) estimator that exploits inter-sample relationships to compute a low-rank Poisson tensor decomposition of the frequency histogram for samples of a multivariate distribution. Our crucial…

Statistics Theory · Mathematics 2026-03-10 Daniel M. Dunlavy , Richard B. Lehoucq , Carolyn D. Mayer , Arvind Prasadan

This paper considers the problem of clustering a collection of unlabeled data points assumed to lie near a union of lower-dimensional planes. As is common in computer vision or unsupervised learning applications, we do not know in advance…

Information Theory · Computer Science 2013-01-31 Mahdi Soltanolkotabi , Emmanuel J. Candés

Zero-inflated datasets, which have an excess of zero outputs, are commonly encountered in problems such as climate or rare event modelling. Conventional machine learning approaches tend to overestimate the non-zeros leading to poor…

Machine Learning · Statistics 2018-03-15 Pashupati Hegde , Markus Heinonen , Samuel Kaski

We put forward a new Bayesian modeling strategy for spatiotemporal count data that enables efficient posterior sampling. Most previous models for such data decompose logarithms of the response Poisson rates into fixed effects and spatial…

Methodology · Statistics 2025-07-29 Yifan Cheng , Cheng Li

Mapping of spatial hotspots, i.e., regions with significantly higher rates of generating cases of certain events (e.g., disease or crime cases), is an important task in diverse societal domains, including public health, public safety,…

Machine Learning · Statistics 2021-10-12 Yiqun Xie , Shashi Shekhar , Yan Li

In recent years, advances in high throughput sequencing technology have led to a need for specialized methods for the analysis of digital gene expression data. While gene expression data measured on a microarray take on continuous values…

Applications · Statistics 2012-02-29 Daniela M. Witten

A new model-based procedure is developed for sparse clustering of functional data that aims to classify a sample of curves into homogeneous groups while jointly detecting the most informative portions of domain. The proposed method is…

Methodology · Statistics 2023-10-04 Fabio Centofanti , Antonio Lepore , Biagio Palumbo

A sparse modeling approach is proposed for analyzing scanning tunneling microscopy topography data, which contains numerous peaks corresponding to surface atoms. The method, based on the relevance vector machine with $\mathrm{L}_1$…

Data Analysis, Statistics and Probability · Physics 2018-03-13 Masamichi J. Miyama , Koji Hukushima

Subspace clustering is the problem of partitioning unlabeled data points into a number of clusters so that data points within one cluster lie approximately on a low-dimensional linear subspace. In many practical scenarios, the…

Machine Learning · Statistics 2019-01-24 Yining Wang , Yu-Xiang Wang , Aarti Singh

We present a hierarchical Bayesian inference approach to estimating the structural properties and the phase space center of a globular cluster (GC) given the spatial and kinematic information of its stars based on lowered isothermal cluster…

Astrophysics of Galaxies · Physics 2023-11-21 Robin Y. Wen , Joshua S. Speagle , Jeremy J. Webb , Gwendolyn M. Eadie

Many data sets cannot be accurately described by standard probability distributions due to the excess number of zero values present. For example, zero-inflation is prevalent in microbiome data and single-cell RNA sequencing data, which…

Methodology · Statistics 2024-11-20 Max Beveridge , Zach Goldstein , Hee Cheol Chung

Imputation of missing values is a strategy for handling non-responses in surveys or data loss in measurement processes, which may be more effective than ignoring them. When the variable represents a count, the literature dealing with this…

Applications · Statistics 2020-07-31 Gilma Hernández-Herrera , Albert Navarro , David Moriña
‹ Prev 1 4 5 6 7 8 10 Next ›