English
Related papers

Related papers: Shape-constrained Estimation of Value Functions

200 papers

This paper presents a numerical method to calculate the value function for a general discounted impulse control problem for piecewise deterministic Markov processes. Our approach is based on a quantization technique for the underlying…

Probability · Mathematics 2011-08-31 Benoîte de Saporta , François Dufour

We consider estimating an expected infinite-horizon cumulative discounted cost/reward contingent on an underlying stochastic process by Monte Carlo simulation. An unbiased estimator based on truncating the cumulative cost at a random…

Numerical Analysis · Mathematics 2020-05-26 Zhenyu Cui , Michael C. Fu , Yijie Peng , Lingjiong Zhu

Vertex direction algorithms have been around for a few decades in the experimental design and mixture models literature. We briefly review this type of algorithm and describe a new member of the family: the support reduction algorithm. The…

Statistics Theory · Mathematics 2009-09-29 Piet Groeneboom , Geurt Jongbloed , Jon A. Wellner

Online nonparametric estimators are gaining popularity due to their efficient computation and competitive generalization abilities. An important example includes variants of stochastic gradient descent. These algorithms often take one…

Statistics Theory · Mathematics 2025-07-08 Tianyu Zhang , Jing Lei

We study a minimax risk of estimating inverse functions on a plane, while keeping an estimator is also invertible. Learning invertibility from data and exploiting an invertible estimator are used in many domains, such as statistics,…

Statistics Theory · Mathematics 2023-12-27 Akifumi Okuno , Masaaki Imaizumi

In this paper, we study parametric nonlinear regression under the Harris recurrent Markov chain framework. We first consider the nonlinear least squares estimators of the parameters in the homoskedastic case, and establish asymptotic theory…

Statistics Theory · Mathematics 2016-09-15 Degui Li , Dag Tjøstheim , Jiti Gao

In this paper, we perform sensitivity analysis for the maximal value function which is the optimal value function for a parametric maximization problem. Our aim is to study various subdifferentials for the maximal value function. We obtain…

Optimization and Control · Mathematics 2023-03-03 L. Guo , J. J. Ye , J. Zhang

We present a numerically efficient approach for learning a risk-neutral measure for paths of simulated spot and option prices up to a finite horizon under convex transaction costs and convex trading constraints. This approach can then be…

Computational Finance · Quantitative Finance 2021-07-15 Hans Buehler , Phillip Murray , Mikko S. Pakkanen , Ben Wood

The main purpose of this chapter is to present some theoretical aspects of parametric estimation of L\'evy processes based on high-frequency sampling, with a focus on infinite activity pure-jump models. Asymptotics for several classes of…

Statistics Theory · Mathematics 2014-09-02 Hiroki Masuda

Reinforcement learning algorithms can solve dynamic decision-making and optimal control problems. With continuous-valued state and input variables, reinforcement learning algorithms must rely on function approximators to represent the value…

Machine Learning · Computer Science 2021-11-16 Jiří Kubalík , Erik Derner , Jan Žegklitz , Robert Babuška

Neural estimators are simulation-based estimators for the parameters of a family of statistical models, which build a direct mapping from the sample to the parameter vector. They benefit from the versatility of available network…

Machine Learning · Statistics 2025-06-24 Almut Rödder , Manuel Hentschel , Sebastian Engelke

We consider a convexity constrained Hamilton-Jacobi-Bellman-type obstacle problem for the value function of a zero-sum differential game with asymmetric information. We propose a convexity-preserving probabilistic numerical scheme for the…

Numerical Analysis · Mathematics 2021-03-26 Ľubomír Baňas , Giorgio Ferrari , Tsiry A. Randrianasolo

We study reinforcement learning in infinite-horizon average-reward settings with linear MDPs. Previous work addresses this problem by approximating the average-reward setting by discounted setting and employing a value iteration-based…

Machine Learning · Computer Science 2025-04-17 Kihyuk Hong , Ambuj Tewari

In a separable Hilbert space, we study the minimization problem of a convex smooth function with Lipschitz continuous gradient whose evaluations are corrupted by random noise. To this end, we associate a stochastic inertial system that…

Optimization and Control · Mathematics 2025-12-18 Chiara Schindler

In this manuscript, we consider a finite nonparametric mixture model with non-independent marginal density functions. Dependence between the marginal densities is modeled using a copula device. Until recently, no deterministic algorithms…

Methodology · Statistics 2025-05-23 Michael Levine

We tackle the problem of estimating risk measures of the infinite-horizon discounted cost within a Markov cost process. The risk measures we study include variance, Value-at-Risk (VaR), and Conditional Value-at-Risk (CVaR). First, we show…

Machine Learning · Computer Science 2024-04-12 Gugan Thoppe , L. A. Prashanth , Sanjay Bhat

Consider the global optimisation of a function $U$ defined on a finite set $V$ endowed with an irreducible and reversible Markov generator.By integration, we extend $U$ to the set $\mathcal{P}(V)$ of probability distributions on $V$ and we…

Functional Analysis · Mathematics 2024-04-16 Laurent Miclo , Nhat-Thang Le

Many real-world decision making tasks require us to choose among several expensive observations. In a sensor network, for example, it is important to select the subset of sensors that is expected to provide the strongest reduction in…

Artificial Intelligence · Computer Science 2014-01-16 Andreas Krause , Carlos Guestrin

This article deals with stochastic processes endowed with the Markov (memoryless) property and evolving over general (uncountable) state spaces. The models further depend on a non-deterministic quantity in the form of a control input, which…

Systems and Control · Computer Science 2015-09-11 Sofie Haesaert , Robert Babuska , Alessandro Abate

This article investigates nonparametric estimation of variance functions for functional data when the mean function is unknown. We obtain asymptotic results for the kernel estimator based on squared residuals. Similar to the finite…

Methodology · Statistics 2008-12-16 Heng Lian
‹ Prev 1 8 9 10 Next ›