中文
相关论文

相关论文: Optimal Unbiased Estimation for Expected Cumulativ…

200 篇论文

In this paper we present an enhancement of the regression-based variance reduction approaches recently proposed in Belomestny et al. This enhancement is based on a truncation of the control variate and allows for a significant reduction of…

概率论 · 数学 2017-11-10 Denis Belomestny , Stefan Häfner , Mikhail Urusov

Policy evaluation via Monte Carlo (MC) simulation is at the core of many MC Reinforcement Learning (RL) algorithms (e.g., policy gradient methods). In this context, the designer of the learning system specifies an interaction budget that…

机器学习 · 计算机科学 2024-10-18 Riccardo Poiani , Nicole Nobili , Alberto Maria Metelli , Marcello Restelli

The most relevant problems in discounted reinforcement learning involve estimating the mean of a function under the stationary distribution of a Markov reward process, such as the expected return in policy evaluation, or the policy gradient…

机器学习 · 计算机科学 2023-04-17 Alberto Maria Metelli , Mirco Mutti , Marcello Restelli

Estimation of value in policy gradient methods is a fundamental problem. Generalized Advantage Estimation (GAE) is an exponentially-weighted estimator of an advantage function similar to $\lambda$-return. It substantially reduces the…

机器学习 · 计算机科学 2023-01-27 Xiulei Song , Yizhao Jin , Greg Slabaugh , Simon Lucas

Some classical uncertainty quantification problems require the estimation of multiple expectations. Estimating all of them accurately is crucial and can have a major impact on the analysis to perform, and standard existing Monte Carlo…

统计方法学 · 统计学 2022-12-02 Julien Demange-Chryst , François Bachoc , Jérôme Morio

This paper introduces a new algorithm for numerically computing equilibrium (i.e. stationary) distributions for Markov chains and Markov jump processes with either a very large finite state space or a countably infinite state space. The…

概率论 · 数学 2022-08-31 Alex Infanger , Peter W. Glynn

Many simulation problems require the estimation of a ratio of two expectations. In recent years Monte Carlo estimators have been proposed that can estimate such ratios without bias. We investigate the theoretical properties of such…

统计理论 · 数学 2019-07-04 Sarat Moka , Dirk P. Kroese , Sandeep Juneja

In classification with a reject option, the classifier is allowed in uncertain cases to abstain from prediction. The classical cost-based model of a reject option classifier requires the cost of rejection to be defined explicitly. An…

机器学习 · 计算机科学 2021-02-01 V. Franc , D. Prusa , V. Voracek

We investigate in this paper an alternative method to simulation based recursive importance sampling procedure to estimate the optimal change of measure for Monte Carlo simulations. We propose an algorithm which combines (vector and…

概率论 · 数学 2011-09-20 Noufel Frikha , Abass Sagna

In this paper we consider the estimation of unknown parameters in Bayesian inverse problems. In most cases of practical interest, there are several barriers to performing such estimation, This includes a numerical approximation of a…

统计方法学 · 统计学 2025-02-07 Neil K. Chada , Ajay Jasra , Mohamed Maama , Raul Tempone

This paper considers an empirical risk minimization problem under heavy-tailed settings, where data does not have finite variance, but only has $p$-th moment with $p \in (1,2)$. Instead of using estimation procedure based on truncated…

机器学习 · 统计学 2023-09-08 Guanhua Fang , Ping Li , Gennady Samorodnitsky

Posterior distributions often feature intractable normalizing constants, called marginal likelihoods or evidence, that are useful for model comparison via Bayes factors. This has motivated a number of methods for estimating ratios of…

统计计算 · 统计学 2018-10-03 Maxime Rischard , Pierre E. Jacob , Natesh Pillai

The expected value of partial perfect information (EVPPI) denotes the value of eliminating uncertainty on a subset of unknown parameters involved in a decision model. The EVPPI can be regarded as a decision-theoretic sensitivity index, and…

统计计算 · 统计学 2016-04-06 Takashi Goda

A weighted Gaussian approximation to tail product-limit process for Pareto-like distributions of randomly right-truncated data is provided and a new consistent and asymptotically normal estimator of the extreme value index is derived. A…

统计理论 · 数学 2015-07-07 Souad Benchaira , Djamel Meraghni , Abdelhakim Necir

We consider selecting the top-$m$ alternatives from a finite number of alternatives via Monte Carlo simulation. Under a Bayesian framework, we formulate the sampling decision as a stochastic dynamic programming problem, and develop a…

最优化与控制 · 数学 2023-08-22 Gongbo Zhang , Yijie Peng , Jianghua Zhang , Enlu Zhou

Parametric stochastic simulators are ubiquitous in science, often featuring high-dimensional input parameters and/or an intractable likelihood. Performing Bayesian parameter inference in this context can be challenging. We present a neural…

机器学习 · 统计学 2021-10-27 Benjamin Kurt Miller , Alex Cole , Patrick Forré , Gilles Louppe , Christoph Weniger

Estimating nested expectations is an important task in computational mathematics and statistics. In this paper we propose a new Monte Carlo method using post-stratification to estimate nested expectations efficiently without taking samples…

数值分析 · 数学 2023-04-28 Tomohiko Hironaka , Takashi Goda

Adaptive Monte Carlo methods are recent variance reduction techniques. In this work, we propose a mathematical setting which greatly relaxes the assumptions needed by for the adaptive importance sampling techniques presented by Vazquez-Abad…

计算金融 · 定量金融 2011-04-28 Bernard Lapeyre , Jérôme Lelong

In Monte Carlo simulations, proposed configurations are accepted or rejected according to an acceptance ratio, which depends on an underlying probability distribution and an a priori sampling probability. By carefully selecting the…

计算物理 · 物理学 2023-02-09 Emanuel Casiano-Diaz , Kipton Barros , Ying Wai Li , Adrian Del Maestro

This paper introduces a new version of the smoothly trimmed mean with a more general version of weights, which can be used as an alternative to the classical trimmed mean. We derive its asymptotic variance and to further investigate its…

统计理论 · 数学 2024-09-10 Elina Kresse , Emils Silins , Janis Valeinis