English
Related papers

Related papers: Hot-spots Detection in Count Data by Poisson Assis…

200 papers

Cluster separation is a task typically tackled by widely used clustering techniques, such as k-means or DBSCAN. However, these algorithms are based on non-perceptual metrics, and our experiments demonstrate that their output does not…

Machine Learning · Computer Science 2025-01-31 Sebastian Hartwig , Christian van Onzenoodt , Dominik Engel , Pedro Hermosilla , Timo Ropinski

We introduce a novel bottom-up approach for the extraction of chart data. Our model utilizes images of charts as inputs and learns to detect keypoints (KP), which are used to reconstruct the components within the plot area. Our novelty lies…

Computer Vision and Pattern Recognition · Computer Science 2023-08-07 Saleem Ahmed , Pengyu Yan , David Doermann , Srirangaraj Setlur , Venu Govindaraju

Motivation: Histone modification constitutes a basic mechanism for the genetic regulation of gene expression. In early 2000s, a powerful technique has emerged that couples chromatin immunoprecipitation with high-throughput sequencing…

Quantitative Methods · Quantitative Biology 2020-12-16 Arnaud Liehrmann , Guillem Rigaill , Toby Dylan Hocking

Scanning devices often produce point clouds exhibiting highly uneven distributions of point samples across the surfaces being captured. Different point cloud subsampling techniques have been proposed to generate more evenly distributed…

Graphics · Computer Science 2023-11-30 Marc Comino-Trinidad , Antonio Chica , Carlos andújar

Count data modeling has been extensively applied in medical sciences to analyze various healthcare datasets. Numerous probability models have been developed to address diverse aspects of healthcare data. In this study, we propose a novel…

Methodology · Statistics 2025-09-04 Peer Bilal Ahmad , Na Elah

State-of-the-art subspace clustering methods are based on self-expressive model, which represents each data point as a linear combination of other data points. By enforcing such representation to be sparse, sparse subspace clustering is…

Machine Learning · Computer Science 2020-05-05 Ying Chen , Chun-Guang Li , Chong You

We describe methods, tools, and a software library called LASPATED, available on GitHub (at https://github.com/vguigues/) to fit models using spatio-temporal data and space-time discretization. A video tutorial for this library is available…

Modeling data with multivariate count responses is a challenging problem due to the discrete nature of the responses. Existing methods for univariate count responses cannot be easily extended to the multivariate case since the dependency…

Methodology · Statistics 2016-08-15 Hao Wu , Xinwei Deng , Naren Ramakrishnan

Atmospheric lidar observations provide a unique capability to directly observe the vertical column of cloud and aerosol scattering properties. Detector and solar background noise, however, hinder the ability of lidar systems to provide…

Optimization and Control · Mathematics 2016-06-21 Willem J. Marais , Robert E. Holz , Yu Hen Hu , Ralph E. Kuehn , Edwin E. Eloranta , Rebecca M. Willett

This article introduces new methods for inference with count data registered on a set of aggregation units. Such data are omnipresent in epidemiology due to confidentiality issues: it is much more common to know the county in which an…

Methodology · Statistics 2017-04-20 Benjamin M. Taylor , Ricardo Andrade-Pacheco , Hugh J. W. Sturrock

We propose a new class of discrete generalized linear models based on the class of Poisson-Tweedie factorial dispersion models with variance of the form $\mu + \phi\mu^p$, where $\mu$ is the mean, $\phi$ and $p$ are the dispersion and…

When modeling the dynamics of infectious disease, the incorporation of contact network information allows for the capture of the non-randomness and heterogeneity of realistic contact patterns. Oftentimes, it is assumed that the underlying…

Populations and Evolution · Quantitative Biology 2024-04-17 Maxwell H. Wang , Jukka-Pekka Onnela

The paper considers a Cox process where the stochastic intensity function for the Poisson data model is itself a non-homogeneous Poisson process. We show that it is possible to obtain the marginal data process, namely a non-homogeneous…

Methodology · Statistics 2023-04-17 Shuying Wang , Stephen G. Walker

Event counts are response variables with non-negative integer values representing the number of times that an event occurs within a fixed domain such as a time interval, a geographical area or a cell of a contingency table. Analysis of…

Point-of-Care Ultrasound (POCUS) refers to clinician-performed and interpreted ultrasonography at the patient's bedside. Interpreting these images requires a high level of expertise, which may not be available during emergencies. In this…

High resolution microarrays and second-generation sequencing platforms are powerful tools to investigate genome-wide alterations in DNA copy number, methylation and gene expression associated with a disease. An integrated genomic profiling…

Applications · Statistics 2013-04-22 Ronglai Shen , Sijian Wang , Qianxing Mo

In this paper we present a method for the unsupervised clustering of high-dimensional binary data, with a special focus on electronic healthcare records. We present a robust and efficient heuristic to face this problem using tensor…

Machine Learning · Statistics 2017-08-31 Matteo Ruffini , Ricard Gavaldà , Esther Limón

Thanks to technological advances leading to near-continuous time observations, emerging multivariate point process data offer new opportunities for causal discovery. However, a key obstacle in achieving this goal is that many relevant…

Machine Learning · Statistics 2021-12-15 Xu Wang , Ali Shojaie

A new two-parameter discrete distribution, namely the PoiG distribution is derived by the convolution of a Poisson variate and an independently distributed geometric random variable. This distribution generalizes both the Poisson and…

Methodology · Statistics 2024-07-11 Anupama Nandi , Subrata Chakraborty , Aniket Biswas

Randomized clinical trials often require large patient cohorts before drawing definitive conclusions, yet abundant observational data from parallel studies remains underutilized due to confounding and hidden biases. To bridge this gap, we…

Machine Learning · Statistics 2025-05-26 Prateek Jaiswal , Esmaeil Keyvanshokooh , Junyu Cao