中文
相关论文

相关论文: Calibrated Tree Priors for Relaxed Phylogenetics a…

200 篇论文

There are several tools available to infer phylogenetic trees, which depict the evolutionary relationships among biological entities such as viral and bacterial strains in infectious outbreaks, or cancerous cells in tumor progression trees.…

数据结构与算法 · 计算机科学 2023-12-22 António Pedro Branco , Cátia Vaz , Alexandre P. Francisco

Confidence calibration assumes a unique ground-truth label per input, yet this assumption fails wherever annotators genuinely disagree. Post-hoc calibrators fitted on majority-voted labels, the standard single-label targets used in…

机器学习 · 计算机科学 2026-03-25 Linwei Tao , Haoyang Luo , Minjing Dong , Chang Xu

Bayesian evidence ratios are widely used to quantify the statistical consistency between different experiments. However, since the evidence ratio is prior dependent, the precise translation between its value and the degree of…

宇宙学与河外天体物理 · 物理学 2021-11-17 V. Miranda , P. Rogozenski , E. Krause

Mixture model-based frameworks are very popular for statistical inference in clustering. While convenient for producing probabilistic estimates of cluster assignments and uncertainty, they are prone to misspecification, which can lead to…

统计理论 · 数学 2026-05-15 Yu Zheng , Leo L. Duan , Arkaprava Roy

Inferring the ancestral state at the root of a phylogenetic tree from states observed at the leaves is a problem arising in evolutionary biology. The simplest technique -- majority rule -- estimates the root state by the most frequently…

种群与进化 · 定量生物学 2014-04-11 Elchanan Mossel , Mike Steel

Decision trees and random forest remain highly competitive for classification on medium-sized, standard datasets due to their robustness, minimal preprocessing requirements, and interpretability. However, a single tree suffers from high…

机器学习 · 统计学 2025-12-02 Cencheng Shen , Yuexiao Dong , Carey E. Priebe

In this paper we introduce objective proper prior distributions for hypothesis testing and model selection based on measures of divergence between the competing models; we call them divergence based (DB) priors. DB priors have simple forms…

统计方法学 · 统计学 2009-02-27 M. J. Bayarri , G. García-Donato

The Yule model and the coalescent model are two neutral stochastic models for generating trees in phylogenetics and population genetics, respectively. Although these models are quite different, they lead to identical distributions…

种群与进化 · 定量生物学 2015-03-17 Sha Zhu , James H. Degnan , Mike Steel

The ratio of two densities provides a direct characterization of their differences. We consider the two-sample comparison problem by estimating this ratio given i.i.d. observations from two distributions. To this end, we propose additive…

统计方法学 · 统计学 2026-04-23 Naoki Awaya , Yuliang Xu , Li Ma

Excellent ranking power along with well calibrated probability estimates are needed in many classification tasks. In this paper, we introduce a technique, Calibrated Boosting-Forest that captures both. This novel technique is an ensemble of…

机器学习 · 统计学 2017-11-15 Haozhen Wu

Forecast systems in science and technology are increasingly moving beyond point prediction toward methods that produce full predictive distributions of future outcomes y, conditional on high-dimensional and complex sequences of inputs x.…

机器学习 · 统计学 2026-03-13 Elizabeth Cucuzzella , Rafael Izbicki , Ann B. Lee

When using complex Bayesian models to combine information, the checking for consistency of the information being combined is good statistical practice. Here a new method is developed for detecting prior-data conflicts in Bayesian models…

统计方法学 · 统计学 2016-11-29 David J. Nott , Xueou Wang , Michael Evans , Berthold-Georg Englert

Ongoing developments in neural network models are continually advancing the state of the art in terms of system accuracy. However, the predicted labels should not be regarded as the only core output; also important is a well-calibrated…

机器学习 · 统计学 2019-01-08 Gil Keren , Nicholas Cummins , Björn Schuller

Bayesian inference has predominantly relied on the Markov chain Monte Carlo (MCMC) algorithm for many years. However, MCMC is computationally laborious, especially for complex phylogenetic models of time trees. This bottleneck has led to…

种群与进化 · 定量生物学 2024-06-27 Mathieu Fourment , Matthew Macaulay , Christiaan J Swanepoel , Xiang Ji , Marc A Suchard , Frederick A Matsen

Many biological studies involve inferring the evolutionary history of a sample of individuals from a large population and interpreting the reconstructed tree. Such an ascertained tree typically represents only a small part of a…

种群与进化 · 定量生物学 2024-08-13 Michael Celentano , William S. DeWitt , Sebastian Prillo , Yun S. Song

A number of recent works have employed decision trees for the construction of explainable partitions that aim to minimize the $k$-means cost function. These works, however, largely ignore metrics related to the depths of the leaves in the…

机器学习 · 计算机科学 2022-08-29 Eduardo Laber , Lucas Murtinho , Felipe Oliveira

Recent studies have demonstrated advantages of information fusion based on sparsity models for multimodal classification. Among several sparsity models, tree-structured sparsity provides a flexible framework for extraction of…

计算机视觉与模式识别 · 计算机科学 2015-02-04 Soheil Bahrampour , Asok Ray , Nasser M. Nasrabadi , Kenneth W. Jenkins

While the performance of machine learning systems has experienced significant improvement in recent years, relatively little attention has been paid to the fundamental question: to what extent can we improve our models? This paper provides…

机器学习 · 计算机科学 2026-05-13 Ryota Ushio , Takashi Ishida , Masashi Sugiyama

Although the analysis of rooted tree shape has wide-ranging applications, notions of tree balance have developed independently in different domains. In computer science, a balanced tree is one that enables efficient updating and retrieval…

定量方法 · 定量生物学 2025-07-14 Veselin Manojlović , Armaan Ahmed , Yannick Viossat , Robert Noble

Algorithms for binary classification based on adaptive tree partitioning are formulated and analyzed for both their risk performance and their friendliness to numerical implementation. The algorithms can be viewed as generating a set…

统计理论 · 数学 2014-11-05 Peter Binev , Albert Cohen , Wolfgang Dahmen , Ronald DeVore