English
Related papers

Related papers: Bayesian significance test for discriminating betw…

200 papers

This study presents a comparative methodological analysis of six machine learning models for survival analysis (MLSA). Using data from nearly 45,000 colorectal cancer patients in the Hospital-Based Cancer Registries of S\~ao Paulo, we…

The difference in restricted mean survival time (RMST) is a clinically meaningful measure to quantify treatment effect in randomized controlled trials, especially when the proportional hazards assumption does not hold. Several frequentist…

We study Bayesian estimation of finite mixture models in a general setup where the number of components is unknown and allowed to grow with the sample size. An assumption on growing number of components is a natural one as the degree of…

Statistics Theory · Mathematics 2022-03-18 Ilsang Ohn , Lizhen Lin

The problem of quickest detection of a change in the distribution of a sequence of independent observations is considered. It is assumed that the pre-change distribution is known (accurately estimated), while the only information about the…

Statistics Theory · Mathematics 2023-09-29 Liyan Xie , Yuchen Liang , Venugopal V. Veeravalli

We consider a stochastic model for species evolution. A new species is born at rate lambda and a species dies at rate mu. A random number, sampled from a given distribution F, is associated with each new species at the time of birth. Every…

Probability · Mathematics 2011-02-15 Herve Guiol , Fabio P. Machado , Rinaldo B. Schinazi

Predictive modelling is vital to guide preventive efforts. Whilst large-scale prospective cohort studies and a diverse toolkit of available machine learning (ML) algorithms have facilitated such survival task efforts, choosing the…

We consider a binary unsupervised classification problem where each observation is associated with an unobserved label that we want to retrieve. More precisely, we assume that there are two groups of observation: normal and abnormal. The…

Machine Learning · Statistics 2011-05-05 Stevenn Volant , Marie-Laure Martin Magniette , Stéphane Robin

Heckman selection model is the most popular econometric model in analysis of data with sample selection. However, selection models with Normal errors cannot accommodate heavy tails in the error distribution. Recently, Marchenko and Genton…

Computation · Statistics 2014-01-08 Peng Ding

In order to cluster or partition data, we often use Expectation-and-Maximization (EM) or Variational approximation with a Gaussian Mixture Model (GMM), which is a parametric probability density function represented as a weighted sum of…

Machine Learning · Computer Science 2013-07-04 Ji Won Yoon

Insights into complex, high-dimensional data can be obtained by discovering features of the data that match or do not match a model of interest. To formalize this task, we introduce the "data selection" problem: finding a lower-dimensional…

Methodology · Statistics 2021-09-10 Eli N. Weinstein , Jeffrey W. Miller

Good large sample performance is typically a minimum requirement of any model selection criterion. This article focuses on the consistency property of the Bayes factor, a commonly used model comparison tool, which has experienced a recent…

Statistics Theory · Mathematics 2016-07-04 Siddhartha Chib , Todd A. Kuffner

Mathematical models support inference and forecasting in ecology and epidemiology, but results depend on the estimation framework. We compare Bayesian and Frequentist approaches across three biological models using four datasets:…

Quantitative Methods · Quantitative Biology 2026-03-31 Mohammed A. Y. Mohammed , Hamed Karami , Gerardo Chowell

Bayesian model selection poses two main challenges: the specification of parameter priors for all models, and the computation of the resulting Bayes factors between models. There is now a large literature on automatic and objective…

Methodology · Statistics 2016-08-11 Leonhard Held , Daniel Sabanés Bové , Isaac Gravestock

In statistics, there are a variety of methods for performing model selection that all stem from slightly different paradigms of statistical inference. The reasons for choosing one particular method over another seem to be based entirely on…

Statistics Theory · Mathematics 2019-01-29 Danica M. Ommen , Christopher P. Saunders

Survival models are used in various fields, such as the development of cancer treatment protocols. Although many statistical and machine learning models have been proposed to achieve accurate survival predictions, little attention has been…

Machine Learning · Computer Science 2020-03-26 Hrushikesh Loya , Pranav Poduval , Deepak Anand , Neeraj Kumar , Amit Sethi

Analysis of matrix-variate data is becoming increasingly common in the literature, particularly in the field of clustering and classification. It is well-known that real data, including real matrix-variate data, often exhibit high levels of…

Methodology · Statistics 2024-07-30 Abbas Mahdavi , Narayanaswamy Balakrishnan , Ahad Jamalizadeh

The multivariate normal linear model is one of the most widely employed models for statistical inference in applied research. Special cases include (multivariate) t testing, (M)AN(C)OVA, (multivariate) multiple regression, and repeated…

Methodology · Statistics 2021-03-15 J. Mulder , H. Hoijtink , X. Gu

In this paper we introduce objective proper prior distributions for hypothesis testing and model selection based on measures of divergence between the competing models; we call them divergence based (DB) priors. DB priors have simple forms…

Methodology · Statistics 2009-02-27 M. J. Bayarri , G. García-Donato

Modeling the diameter distribution of trees in forest stands is a common forestry task that supports key biologically and economically relevant management decisions. The choice of model used to represent the diameter distribution and how to…

Applications · Statistics 2019-11-26 Mahdi Teimouri , Jeffrey W. Doser , Andrew O. Finley

Testing differences between a treatment and control group is common practice in biomedical research like randomized controlled trials (RCT). The standard two-sample t-test relies on null hypothesis significance testing (NHST) via p-values,…

Methodology · Statistics 2020-05-18 Riko Kelter