English
Related papers

Related papers: A nonparametric HMM for genetic imputation and coa…

200 papers

Hidden Markov jump processes are an attractive approach for modeling clinical disease progression data because they are explainable and capable of handling both irregularly sampled and noisy data. Most applications in this context consider…

Methodology · Statistics 2019-10-15 Rui Meng , Soper Braden , Jan Nygard , Mari Nygrad , Herbert Lee

We describe a generalization of the Hierarchical Dirichlet Process Hidden Markov Model (HDP-HMM) which is able to encode prior information that state transitions are more likely between "nearby" states. This is accomplished by defining a…

Machine Learning · Statistics 2017-07-24 Colin Reimer Dawson , Chaofan Huang , Clayton T. Morrison

Standard practice in Hidden Markov Model (HMM) selection favors the candidate with the highest full-sequence likelihood, although this is equivalent to making a decision based on a single realization. We introduce a \emph{fragment-based}…

Methodology · Statistics 2025-05-01 Carlos M. Hernandez-Suarez , Osval A. Montesinos-López

Suppose that we are given a time series where consecutive samples are believed to come from a probabilistic source, that the source changes from time to time and that the total number of sources is fixed. Our objective is to estimate the…

Information Theory · Computer Science 2018-04-24 Mark Kozdoba , Shie Mannor

Datasets containing large samples of time-to-event data arising from several small heterogeneous groups are commonly encountered in statistics. This presents problems as they cannot be pooled directly due to their heterogeneity or analyzed…

Machine Learning · Statistics 2016-12-05 Alexandre Piché , Russell Steele , Ian Shrier , Stephanie Long

We introduce a nonparametric model for inferring time-evolving, unobserved probability distributions from discrete-time data consisting of unlabelled partitions. The latent process is a two-parameter Poisson-Dirichlet diffusion, and…

Methodology · Statistics 2026-05-19 Marco Dalla Pria , Matteo Ruggiero , Dario Spanò

Sequence analysis is being more and more widely used for the analysis of social sequences and other multivariate categorical time series data. However, it is often complex to describe, visualize, and compare large sequence data, especially…

Computation · Statistics 2021-03-22 Satu Helske , Jouni Helske

Labeling of sequential data is a prevalent meta-problem for a wide range of real world applications. While the first-order Hidden Markov Models (HMM) provides a fundamental approach for unsupervised sequential labeling, the basic model does…

Machine Learning · Computer Science 2019-04-08 Maoying Qiao , Wei Bian , Richard Yida Xu , Dacheng Tao

The hidden Markov model (HMM) has been a workhorse of single molecule data analysis and is now commonly used as a standalone tool in time series analysis or in conjunction with other analyses methods such as tracking. Here we provide a…

Data Analysis, Statistics and Probability · Physics 2017-06-28 Ioannis Sgouralis , Steve Presse

Gene-gene and gene-environment interactions are widely believed to play significant roles in explaining the variability of complex traits. While substantial research exists in this area, a comprehensive statistical framework that addresses…

Methodology · Statistics 2026-02-18 Durba Bhattacharya , Sourabh Bhattacharya

A Hidden Markov Model (HMM) is a common statistical model which is widely used for analysis of biological sequence data and other sequential phenomena. In the present paper we show how HMMs can be extended with side-constraints and present…

Artificial Intelligence · Computer Science 2010-08-02 Henning Christiansen , Christian Theil Have , Ole Torp Lassen , Matthieu Petit

We present a new algorithm for discovering patterns in time series and other sequential data. We exhibit a reliable procedure for building the minimal set of hidden, Markovian states that is statistically capable of producing the behavior…

Machine Learning · Computer Science 2007-05-23 Cosma Rohilla Shalizi , Kristina Lisa Shalizi , James P. Crutchfield

A number of statistical models have been successfully developed for the analysis of high-throughput data from a single source, but few methods are available for integrating data from different sources. Here we focus on integrating gene…

Hidden Markov models (HMMs) have been extensively used in the univariate and multivariate literature. However, there has been an increased interest in the analysis of matrix-variate data over the recent years. In this manuscript we…

Methodology · Statistics 2021-07-16 Salvatore D. Tomarchio , Antonio Punzo , Antonello Maruotti

This paper proposes a new Bayesian multiple change-point model which is based on the hidden Markov approach. The Dirichlet process hidden Markov model does not require the specification of the number of change-points a priori. Hence our…

Statistics Theory · Mathematics 2015-05-08 Stanley I. M. Ko , Terence T. L. Chong , Pulak Ghosh

The development of coalescent theory paved the way to statistical inference from population genetic data. In the genomic era, however, coalescent models are limited due to the complexity of the underlying ancestral recombination graph. The…

Populations and Evolution · Quantitative Biology 2021-06-30 Julien Y. Dutheil

The technological applications of hidden Markov models have been extremely diverse and successful, including natural language processing, gesture recognition, gene sequencing, and Kalman filtering of physical measurements. HMMs are highly…

Algebraic Geometry · Mathematics 2012-09-04 Andrew J. Critch

Traditional hidden Markov models have been a useful tool to understand and model stochastic dynamic data; in the case of non-Gaussian data, models such as mixture of Gaussian hidden Markov models can be used. However, these suffer from the…

Machine Learning · Statistics 2023-05-16 Carlos Puerto-Santana , Concha Bielza , Pedro Larrañaga , Gustav Eje Henter

Chromosomal DNA is characterized by variation between individuals at the level of entire chromosomes (e.g., aneuploidy in which the chromosome copy number is altered), segmental changes (including insertions, deletions, inversions, and…

Applications · Statistics 2008-07-30 Robert B. Scharpf , Giovanni Parmigiani , Jonathan Pevsner , Ingo Ruczinski

Clustering is one of the most widely used procedures in the analysis of microarray data, for example with the goal of discovering cancer subtypes based on observed heterogeneity of genetic marks between different tissues. It is well-known…

Methodology · Statistics 2009-04-21 Heng Lian