English
Related papers

Related papers: Stochastic Policy Gradient Methods in the Uncertai…

200 papers

We propose a method for finding approximate compilations of quantum unitary transformations, based on techniques from policy gradient reinforcement learning. The choice of a stochastic policy allows us to rephrase the optimization problem…

Quantum Physics · Physics 2022-09-14 David A. Herrera-Martí

This study provides a consistent and efficient pricing method for both Standard & Poor's 500 Index (SPX) options and the Chicago Board Options Exchange's Volatility Index (VIX) options under a multiscale stochastic volatility model. To…

Mathematical Finance · Quantitative Finance 2019-09-24 Jaegi Jeon , Geonwoo Kim , Jeonggyu Huh

Power systems that need to integrate renewables at a large scale must account for the high levels of uncertainty introduced by these power sources. This can be accomplished with a system of many distributed grid-level storage devices.…

Optimization and Control · Mathematics 2020-02-04 Joseph L. Durante , Juliana Nascimento , Warren B. Powell

In many sequential decision-making problems we may want to manage risk by minimizing some measure of variability in rewards in addition to maximizing a standard criterion. Variance related risk measures are among the most common…

Machine Learning · Computer Science 2015-03-19 Prashanth L. A. , Mohammad Ghavamzadeh

Non-equilibrium phenomena occur not only in physical world, but also in finance. In this work, stochastic relaxational dynamics (together with path integrals) is applied to option pricing theory. A recently proposed model (by Ilinski et…

Statistical Mechanics · Physics 2009-10-31 Matthias Otto

We consider multistage stochastic linear optimization problems combining joint dynamic probabilistic constraints with hard constraints. We develop a method for projecting decision rules onto hard constraints of wait-and-see type. We…

Optimization and Control · Mathematics 2016-09-16 Vincent Guigues , Rene Henrion

This paper is concerned with portfolio optimization models for creating high-quality lists of recommended items to balance the accuracy and diversity of recommendations. However, the statistics (i.e., expectation and covariance of ratings)…

Information Retrieval · Computer Science 2024-10-01 Tomoya Yanagi , Shunnosuke Ikeda , Yuichi Takano

Policy gradient methods, which have been extensively studied in the last decade, offer an effective and efficient framework for reinforcement learning problems. However, their performances can often be unsatisfactory, suffering from…

Machine Learning · Computer Science 2026-01-27 Shihab Ahmed , El Houcine Bergou , Aritra Dutta , Yue Wang

We derive high-order compact finite difference schemes for option pricing in stochastic volatility models on non-uniform grids. The schemes are fourth-order accurate in space and second-order accurate in time for vanishing correlation. In…

Computational Finance · Quantitative Finance 2014-05-12 Bertram Düring , Michel Fournié , Christof Heuer

This paper presents a robust version of the stratified sampling method when multiple uncertain input models are considered for stochastic simulation. Various variance reduction techniques have demonstrated their superior performance in…

Optimization and Control · Mathematics 2023-06-16 Seung Min Baik , Eunshin Byon , Young Myoung Ko

Recent years have seen an increased level of interest in pricing equity options under a stochastic volatility model such as the Heston model. Often, simulating a Heston model is difficult, as a standard finite difference scheme may lead to…

Computational Finance · Quantitative Finance 2011-11-28 Ian Iscoe , Asif Lakhany

We consider a dynamic portfolio optimization problem that incorporates predictable returns, instantaneous transaction costs, price impact, and stochastic volatility, extending the classical results of Garleanu and Pedersen (2013), which…

Computational Finance · Quantitative Finance 2025-07-24 Patrick Chan , Ronnie Sircar , Iosif Zimbidis

Domain randomization is a simple, effective, and flexible scheme for obtaining robust feedback policies aimed at reducing the sim-to-real gap due to model mismatch. While domain randomization methods have yielded impressive demonstrations…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Alex Nguyen-Le , Nikolai Matni

The ability to accurately predict human behavior is central to the safety and efficiency of robot autonomy in interactive settings. Unfortunately, robots often lack access to key information on which these predictions may hinge, such as…

Robotics · Computer Science 2022-06-07 Haimin Hu , Jaime F. Fisac

We consider stochastic model predictive control of a multi-agent systems with constraints on the probabilities of inter-agent collisions. We first study a sample-based approximation of the collision probabilities and use this approximation…

Systems and Control · Computer Science 2011-08-17 Daniel Lyons , Jan-P. Calliess , Uwe D. Hanebeck

Based on criteria of mathematical simplicity and consistency with empirical market data, a stochastic volatility model is constructed, the volatility process being driven by fractional noise. Price return statistics and asymptotic behavior…

Probability · Mathematics 2008-12-02 Rui Vilela Mendes , M. J. Oliveira

Stabilizing an unknown control system is one of the most fundamental problems in control systems engineering. In this paper, we provide a simple, model-free algorithm for stabilizing fully observed dynamical systems. While model-free…

Systems and Control · Electrical Eng. & Systems 2021-10-14 Juan C. Perdomo , Jack Umenberger , Max Simchowitz

While the optimization landscape of policy gradient methods has been recently investigated for partially observed linear systems in terms of both static output feedback and dynamical controllers, they only provide convergence guarantees to…

Optimization and Control · Mathematics 2023-04-25 Feiran Zhao , Xingyun Fu , Keyou You

Continuous-time Markov decision processes are an important class of models in a wide range of applications, ranging from cyber-physical systems to synthetic biology. A central problem is how to devise a policy to control the system in order…

Systems and Control · Computer Science 2016-06-01 Ezio Bartocci , Luca Bortolussi , Tomǎš Brázdil , Dimitrios Milios , Guido Sanguinetti

Stochastic differential equation (SDE) models are the foundation for pricing and hedging financial derivatives. The drift and volatility functions in SDE models are typically chosen to be algebraic functions with a small number (less than…

Computational Finance · Quantitative Finance 2024-06-04 Lei Fan , Justin Sirignano