English
Related papers

Related papers: Ignorable and non-ignorable missing data in hidden…

200 papers

We introduce a self-censoring model for multivariate nonignorable nonmonotone missing data, where the missingness process of each outcome is affected by its own value and is associated with missingness indicators of other outcomes, while…

Methodology · Statistics 2022-10-03 Yilin Li , Wang Miao , Ilya Shpitser , Eric J. Tchetgen Tchetgen

Trial-based cost-effectiveness analyses (CEAs) are an important source of evidence in the assessment of health interventions. In these studies, cost and effectiveness outcomes are commonly measured at multiple time points, but some…

Methodology · Statistics 2022-03-30 Andrea Gabrio , Catrin Plumpton , Sube Banerjee , Baptiste Leurent

Researchers regularly perform conditional prediction using imputed values of missing data. However, applications of imputation often lack a firm foundation in statistical theory. This paper originated when we were unable to find analysis…

Econometrics · Economics 2021-02-24 Charles F Manski , Michael Gmeiner , Anat Tamburc

Forecasting tasks using large datasets gathering thousands of heterogeneous time series is a crucial statistical problem in numerous sectors. The main challenge is to model a rich variety of time series, leverage any available external…

Machine Learning · Computer Science 2024-04-18 Etienne David , Jean Bellot , Sylvain Le Corff

Markov state models (MSMs) are widely employed to analyze the kinetics of complex systems. But despite their effectiveness in many applications, MSMs are prone to systematic or statistical errors, often exacerbated by suboptimal…

Data Analysis, Statistics and Probability · Physics 2025-08-12 Yehor Tuchkov , Luke Evans , Sonya M. Hanson , Erik H. Thiede

We develop a flexible spline-based Bayesian hidden Markov model stochastic weather generator to statistically model daily precipitation over time by season at individual locations. The model naturally accounts for missing data (considered…

Applications · Statistics 2022-07-19 Christopher J. Paciorek

Multiple imputation is a well-established general technique for analyzing data with missing values. A convenient way to implement multiple imputation is sequential regression multiple imputation (SRMI), also called chained equations…

Multi-state capture-recapture data comprise individual-specific sighting histories together with information on individuals' states related, for example, to breeding status, infection level, or geographical location. Such data are often…

Applications · Statistics 2023-11-16 Sina Mews , Roland Langrock , Ruth King , Nicola Quick

State-space models effectively model multivariate time series by updating over time a representation of the system state from which predictions are made. The state representation is usually a vector without any explicit structure.…

Machine Learning · Computer Science 2026-04-07 Daniele Zambon , Andrea Cini , Cesare Alippi

We propose an inferential approach for maximum likelihood estimation of the hidden Markov models for continuous responses. We extend to the case of longitudinal observations the finite mixture model of multivariate Gaussian distributions…

Methodology · Statistics 2021-07-01 Silvia Pandolfi , Francesco Bartolucci , Fulvia Pennoni

The Hidden Markov Model (HMM) is one of the most widely used statistical models for sequential data analysis. One of the key reasons for this versatility is the ability of HMM to deal with missing data. However, standard HMM learning…

Machine Learning · Statistics 2023-07-04 Binyamin Perets , Mark Kozdoba , Shie Mannor

In the traditional framework of spectral learning of stochastic time series models, model parameters are estimated based on trajectories of fully recorded observations. However, real-world time series data often contain missing values, and…

Machine Learning · Computer Science 2018-10-22 Tianlin Liu

Missing data are ubiquitous in many domains including healthcare. When these data entries are not missing completely at random, the (conditional) independence relations in the observed data may be different from those in the complete data…

Machine Learning · Computer Science 2020-07-14 Ruibo Tu , Kun Zhang , Paul Ackermann , Bo Christer Bertilson , Clark Glymour , Hedvig Kjellström , Cheng Zhang

This work aims at providing a new model for time series classification based on learning from just one example. We assume that time series can be well characterized as a parametric random process, a sort of Hidden semi-Markov Model…

Machine Learning · Statistics 2022-11-18 Adrián Pérez Herrero , Paulo Félix Lamas , Jesús María Rodríguez Presedo

Time-series models typically assume untainted and legitimate streams of data. However, a self-interested adversary may have incentive to corrupt this data, thereby altering a decision maker's inference. Within the broader field of…

Cryptography and Security · Computer Science 2024-02-22 William N. Caballero , Jose Manuel Camacho , Tahir Ekin , Roi Naveiro

Dealing with missing data poses significant challenges in predictive analysis, often leading to biased conclusions when oversimplified assumptions about the missing data process are made. In cases where the data are missing not at random…

Methodology · Statistics 2024-12-20 Yong Chen Goh , Wuu Kuang Soh , Andrew C. Parnell , Keefe Murphy

Environmental time series data observed at high frequencies can be studied with approaches such as hidden Markov and semi-Markov models (HMM and HSMM). HSMMs extend the HMM by explicitly modeling the time spent in each state. In a…

Multimorbidity in older adults is common, heterogeneous, and highly dynamic, and it is strongly associated with disability and increased healthcare utilization. However, existing approaches to studying multimorbidity trajectories are…

Nonmonotone missing data is a common problem in scientific studies. The conventional ignorability and missing-at-random (MAR) conditions are unlikely to hold for nonmonotone missing data and data analysis can be very challenging with few…

Methodology · Statistics 2022-07-07 Gang Cheng , Yen-Chi Chen , Maureen A. Smith , Ying-Qi Zhao

Identifiability of parameters is an essential property for a statistical model to be useful in most settings. However, establishing parameter identifiability for Bayesian networks with hidden variables remains challenging. In the context of…

Statistics Theory · Mathematics 2014-06-04 Elizabeth S. Allman , John A. Rhodes , Elena Stanghellini , Marco Valtorta