English
Related papers

Related papers: Rao-Blackwellizing the Straight-Through Gumbel-Sof…

200 papers

Recent progress in deep latent variable models has largely been driven by the development of flexible and scalable variational inference methods. Variational training of this type involves maximizing a lower bound on the log-likelihood,…

Machine Learning · Computer Science 2016-06-02 Andriy Mnih , Danilo J. Rezende

Within many machine learning algorithms, a fundamental problem concerns efficient calculation of an unbiased gradient wrt parameters $\gammav$ for expectation-based objectives $\Ebb_{q_{\gammav} (\yv)} [f(\yv)]$. Most existing methods…

Machine Learning · Statistics 2019-01-21 Yulai Cong , Miaoyun Zhao , Ke Bai , Lawrence Carin

To address the challenge of backpropagating the gradient through categorical variables, we propose the augment-REINFORCE-swap-merge (ARSM) gradient estimator that is unbiased and has low variance. ARSM first uses variable augmentation,…

Machine Learning · Statistics 2019-12-24 Mingzhang Yin , Yuguang Yue , Mingyuan Zhou

We present a filtering framework for online joint state estimation and parameter identification in nonlinear, time-varying systems. The algorithm uses Rao-Blackwellization technique to infer joint state-parameter posteriors efficiently. In…

Systems and Control · Electrical Eng. & Systems 2026-03-25 Milad Banitalebi Dehkordi , Manas Mejari , Dario Piga

Recent variational inference methods use stochastic gradient estimators whose variance is not well understood. Theoretical guarantees for these estimators are important to understand when these methods will or will not work. This paper…

Machine Learning · Computer Science 2019-10-29 Justin Domke

We develop two new estimators for a general class of stationary GARCH models with possibly heavy tailed asymmetrically distributed errors, covering processes with symmetric and asymmetric feedback like GARCH, Asymmetric GARCH, VGARCH and…

Statistics Theory · Mathematics 2015-07-29 Jonathan B. Hill

Recently, latent reasoning has been introduced into large language models (LLMs) to leverage rich information within a continuous space. However, without stochastic sampling, these methods inevitably collapse to deterministic inference,…

Machine Learning · Computer Science 2026-05-12 Yuyan Zhou , Jiarui Yu , Hande Dong , Zhezheng Hao , Hong Wang , Jianqing Zhang , Qiang Lin

Inference on vertex-aligned graphs is of wide theoretical and practical importance.There are, however, few flexible and tractable statistical models for correlated graphs, and even fewer comprehensive approaches to parametric inference on…

Statistics Theory · Mathematics 2021-04-01 Donniell E. Fishkind , Avanti Athreya , Lingyao Meng , Vince Lyzinski , Carey E. Priebe

We present an in-depth analysis of the sources of variance in state-of-the-art unbiased volumetric transmittance estimators, and propose several new methods for improving their efficiency. These combine to produce a single estimator that is…

Graphics · Computer Science 2021-08-20 Markus Kettunen , Eugene d'Eon , Jacopo Pantaleoni , Jan Novak

Sequential Monte Carlo (SMC) methods, such as the particle filter, are by now one of the standard computational techniques for addressing the filtering problem in general state-space models. However, many applications require…

Computation · Statistics 2016-04-20 Fredrik Lindsten , Pete Bunch , Simo Särkkä , Thomas B. Schön , Simon J. Godsill

In system identification, estimating parameters of a model using limited observations results in poor identifiability. To cope with this issue, we propose a new method to simultaneously select and estimate sensitive parameters as key model…

The mean dimension of a black box function of $d$ variables is a convenient way to summarize the extent to which it is dominated by high or low order interactions. It is expressed in terms of $2^d-1$ variance components but it can be…

Methodology · Statistics 2021-01-01 Christopher Hoyt , Art B. Owen

This paper develops a flexible method for decreasing the variance of estimators for complex experiment effect metrics (e.g. ratio metrics) while retaining asymptotic unbiasedness. This method uses the auxiliary information about the…

Statistics Theory · Mathematics 2019-04-09 Reza Hosseini , Amir Najmi

This study explores the application of a two-level algorithm to enhance the signal-to-noise ratio of glueball calculations in four-dimensional $\mathrm{SU(3)}$ pure gauge theory. Our findings demonstrate that the statistical errors exhibit…

High Energy Physics - Lattice · Physics 2024-06-19 Lorenzo Barca , Francesco Knechtli , Sofie Martins , Michael Peardon , Stefan Schaefer , Juan Andrés Urrea-Niño

Countless signal processing applications include the reconstruction of signals from few indirect linear measurements. The design of effective measurement operators is typically constrained by the underlying hardware and physics, posing a…

Machine Learning · Computer Science 2023-05-23 Jonathan Sauder , Martin Genzel , Peter Jung

Variance estimation in the linear model when $p > n$ is a difficult problem. Standard least squares estimation techniques do not apply. Several variance estimators have been proposed in the literature, all with accompanying asymptotic…

Methodology · Statistics 2014-01-30 Stephen Reid , Robert Tibshirani , Jerome Friedman

In optimizing real-world structures, due to fabrication or budgetary restraints, the design variables may be restricted to a set of standard engineering choices. Such variables, commonly called categorical variables, are discrete and…

Computational Engineering, Finance, and Science · Computer Science 2025-01-03 Mehran Ebrahimi , Hyunmin Cheong , Pradeep Kumar Jayaraman , Farhad Javid

Reinforcement learning with verifiable rewards (RLVR) is effective for training large language models on deterministic outcome reasoning tasks. Prior work shows RLVR works with few prompts, but prompt selection is often based only on…

Machine Learning · Computer Science 2026-05-07 Yujuan Pang , Jiaxin Li , Xin Sheng , Ran Peng , Yong Ma

Population Monte Carlo has been introduced as a sequential importance sampling technique to overcome poor fit of the importance function. In this paper, we compare the performances of the original Population Monte Carlo algorithm with a…

Computation · Statistics 2008-02-26 Alessandra Iacobucci , Jean-Michel Marin , Christian Robert

We deal with the problem of gradient estimation for stochastic differentiable relaxations of algorithms, operators, simulators, and other non-differentiable functions. Stochastic smoothing conventionally perturbs the input of a…

Machine Learning · Computer Science 2024-10-11 Felix Petersen , Christian Borgelt , Aashwin Mishra , Stefano Ermon