中文
相关论文

相关论文: The fine print on tempered posteriors

200 篇论文

Recent work has observed that one can outperform exact inference in Bayesian neural networks by tuning the "temperature" of the posterior on a validation set (the "cold posterior" effect). To help interpret this phenomenon, we argue that…

机器学习 · 统计学 2020-08-04 Ben Adlam , Jasper Snoek , Samuel L. Smith

Bayesian methods feature useful properties for solving inverse problems, such as tomographic reconstruction. The prior distribution introduces regularization, which helps solving the ill-posed problem and reduces overfitting. In practice,…

图像与视频处理 · 电气工程与系统科学 2021-12-02 Max-Heinrich Laves , Malte Tölle , Alexander Schlaefer , Sandy Engelhardt

We investigate the cold posterior effect through the lens of PAC-Bayes generalization bounds. We argue that in the non-asymptotic setting, when the number of training samples is (relatively) small, discussions of the cold posterior effect…

机器学习 · 计算机科学 2022-06-23 Konstantinos Pitas , Julyan Arbel

In recent years, inconsistency in Bayesian deep learning has attracted significant attention. Tempered or generalized posterior distributions are frequently employed as direct and effective solutions. Nonetheless, the underlying mechanisms…

机器学习 · 计算机科学 2025-09-23 Yinsong Chen , Samson S. Yu , Zhong Li , Chee Peng Lim

Model-based filtering is often carried out while subject to an imperfect model, as learning partially-observable stochastic systems remains a challenge. Recent work on Bayesian inference found that tempering the likelihood or full posterior…

系统与控制 · 电气工程与系统科学 2025-12-03 Menno van Zutphen , Domagoj Herceg , Giannis Delimpaltadakis , Duarte J. Antunes

Aleatoric uncertainty captures the inherent randomness of the data, such as measurement noise. In Bayesian regression, we often use a Gaussian observation model, where we control the level of aleatoric uncertainty with a noise variance…

机器学习 · 计算机科学 2022-03-31 Sanyam Kapoor , Wesley J. Maddox , Pavel Izmailov , Andrew Gordon Wilson

Bayesian inference provides a principled probabilistic framework for quantifying uncertainty by updating beliefs based on prior knowledge and observed data through Bayes' theorem. In Bayesian deep learning, neural network weights are…

机器学习 · 计算机科学 2024-10-22 Yijie Zhang

The Laplace approximation is a popular method for constructing a Gaussian approximation to the Bayesian posterior and thereby approximating the posterior mean and variance. But approximation quality is a concern. One might consider using…

统计理论 · 数学 2025-06-17 Mikołaj J. Kasprzak , Ryan Giordano , Tamara Broderick

Posterior tempering reduces the influence of the likelihood in the calculation of the posterior by raising the likelihood to a fractional power $\alpha$. The resulting power posterior - also known as an $\alpha$-posterior or fractional…

统计理论 · 数学 2026-01-15 Ruchira Ray , Marco Avella Medina , Cynthia Rush

To get Bayesian neural networks to perform comparably to standard neural networks it is usually necessary to artificially reduce uncertainty using a "tempered" or "cold" posterior. This is extremely concerning: if the prior is accurate,…

机器学习 · 统计学 2021-04-28 Laurence Aitchison

While Bayesian neural networks (BNNs) provide a sound and principled alternative to standard neural networks, an artificial sharpening of the posterior usually needs to be applied to reach comparable performance. This is in stark contrast…

机器学习 · 计算机科学 2023-07-19 Gregor Bachmann , Lorenzo Noci , Thomas Hofmann

We present a simple method to obtain optimal posterior distributions and improve the quality of Bayesian inference with reduced human and computational effort. Bayes' Theorem is reformulated in the language of statistical mechanics, wherein…

统计方法学 · 统计学 2026-04-28 Alfred C. K. Farris

Benchmark datasets used for image classification tend to have very low levels of label noise. When Bayesian neural networks are trained on these datasets, they often underfit, misrepresenting the aleatoric uncertainty of the data. A common…

机器学习 · 计算机科学 2024-03-05 Martin Marek , Brooks Paige , Pavel Izmailov

The Cold Posterior Effect (CPE) is a phenomenon in Bayesian Deep Learning (BDL), where tempering the posterior to a cold temperature often improves the predictive performance of the posterior predictive distribution (PPD). Although the term…

机器学习 · 统计学 2025-10-27 Kenyon Ng , Chris van der Heide , Liam Hodgkinson , Susan Wei

PAC-Bayesian algorithms and Gibbs posteriors are gaining popularity due to their robustness against model misspecification even when Bayesian inference is inconsistent. The PAC-Bayesian alpha-posterior is a generalization of the standard…

机器学习 · 计算机科学 2020-04-23 Lucie Perrotta

The Laplace approximation (LA) to posteriors is a ubiquitous tool to simplify Bayesian computation, particularly in the high-dimensional settings arising in Bayesian inverse problems. Precisely quantifying the LA accuracy is a challenging…

统计理论 · 数学 2025-09-10 Anya Katsevich , Vladimir Spokoiny

Bayesian optimization (BO) iteratively fits a Gaussian process (GP) surrogate to accumulated evaluations and selects new queries via an acquisition function such as expected improvement (EI). In practice, BO often concentrates evaluations…

统计方法学 · 统计学 2026-01-13 Jiguang Li , Hengrui Luo

Uncertainty quantification in reinforcement learning can greatly improve exploration and robustness. Approximate Bayesian approaches have recently been popularized to quantify uncertainty in model-free algorithms. However, so far the focus…

机器学习 · 计算机科学 2025-09-01 Pascal R. van der Vaart , Neil Yorke-Smith , Matthijs T. J. Spaan

The cold posterior effect (CPE) (Wenzel et al., 2020) in Bayesian deep learning shows that, for posteriors with a temperature $T<1$, the resulting posterior predictive could have better performances than the Bayesian posterior ($T=1$). As…

机器学习 · 统计学 2023-10-03 Yijie Zhang , Yi-Shan Wu , Luis A. Ortega , Andrés R. Masegosa

Temperature scaling is a simple method that allows to control the uncertainty of probabilistic models. It is mostly used in two contexts: improving the calibration of classifiers and tuning the stochasticity of large language models (LLMs).…

机器学习 · 统计学 2026-05-28 Pierre-Alexandre Mattei , Bruno Loureiro
‹ 上一页 1 2 3 10 下一页 ›