English
Related papers

Related papers: Convergence of Proximal Policy Gradient Method for…

200 papers

Existing deterministic variational inference approaches for diffusion processes use simple proposals and target the marginal density of the posterior. We construct the variational process as a controlled version of the prior process and…

Machine Learning · Computer Science 2021-03-02 Christian Wildner , Heinz Koeppl

In this work, we study the gradient projection method for solving a class of stochastic control problems by using a mesh free approximation approach to implement spatial dimension approximation. Our main contribution is to extend the…

Optimization and Control · Mathematics 2021-04-21 Hui Sun , Feng Bao

We consider a control problem where the system is driven by a decoupled as well as a coupled forward-backward stochastic differential equation. We prove the existence of an optimal control in the class of relaxed controls, which are…

Optimization and Control · Mathematics 2017-01-31 Fouzia Baghery , Nabil Khelfallah , Brahim Mezerdi , Isabelle Turpin

Model-based policy optimization is a well-established framework for designing reliable and high-performance controllers across a wide range of control applications. Recently, this approach has been extended to model predictive control…

Systems and Control · Electrical Eng. & Systems 2026-04-15 Riccardo Zuliani , Efe C. Balta , John Lygeros

We investigate a stochastic optimal control problem where the controlled system is depicted as a stochastic differential delayed equation; however, at the terminal time, the state is constrained in a convex set. We firstly introduce an…

Probability · Mathematics 2017-05-12 Jiaqiang Wen , Yufeng Shi

In this article, we discuss two algorithms tailored to discrete-time deterministic finite-horizon nonlinear optimal control problems or so-called deterministic trajectory optimization problems. Both algorithms can be derived from an…

Optimization and Control · Mathematics 2024-12-10 Mohammad Mahmoudi Filabadi , Tom Lefebvre , Guillaume Crevecoeur

Optimal control under uncertainty is a prevailing challenge for many reasons. One of the critical difficulties lies in producing tractable solutions for the underlying stochastic optimization problem. We show how advanced approximate…

Machine Learning · Computer Science 2024-10-28 Joe Watson , Hany Abdulsamad , Rolf Findeisen , Jan Peters

We develop model-based methods for solving stochastic convex optimization problems, introducing the approximate-proximal point, or aProx, family, which includes stochastic subgradient, proximal point, and bundle methods. When the modeling…

Optimization and Control · Mathematics 2019-09-20 Hilal Asi , John C. Duchi

We consider control-constrained linear-quadratic optimal control problems on evolving surfaces. In order to formulate well-posed problems, we prove existence and uniqueness of weak solutions for the state equation, in the sense of…

Optimization and Control · Mathematics 2015-03-19 Morten Vierling

The inverse problem of backward diffusion is known to be ill-posed and highly unstable. Backward diffusion processes appear naturally in image enhancement and deblurring applications. It is therefore greatly desirable to establish a…

Numerical Analysis · Mathematics 2020-06-18 Leif Bergerhoff , Marcelo Cárdenas , Joachim Weickert , Martin Welk

Driven by the need to solve increasingly complex optimization problems in signal processing and machine learning, there has been increasing interest in understanding the behavior of gradient-descent algorithms in non-convex environments.…

Optimization and Control · Mathematics 2019-07-04 Stefan Vlaski , Ali H. Sayed

In this paper, we study a stochastic optimal control problem under a type of consistent convex expectation dominated by G-expectation. By the separation theorem for convex sets, we get the representation theorems for this convex expectation…

Optimization and Control · Mathematics 2024-08-21 Xiaojuan Li , Mingshang Hu

Sparse learning is a very important tool for mining useful information and patterns from high dimensional data. Non-convex non-smooth regularized learning problems play essential roles in sparse learning, and have drawn extensive attentions…

Machine Learning · Computer Science 2020-10-22 Guannan Liang , Qianqian Tong , Jiahao Ding , Miao Pan , Jinbo Bi

In this paper, a gradient-free distributed algorithm is introduced to solve a set constrained optimization problem under a directed communication network. Specifically, at each time-step, the agents locally compute a so-called…

Optimization and Control · Mathematics 2021-09-06 Yipeng Pang , Guoqiang Hu

A new method for stochastic control based on neural networks and using randomisation of discrete random variables is proposed and applied to optimal stopping time problems. The method models directly the policy and does not need the…

Computational Finance · Quantitative Finance 2021-01-11 Thomas Deschatre , Joseph Mikael

We analyze stochastic algorithms for optimizing nonconvex, nonsmooth finite-sum problems, where the nonconvex part is smooth and the nonsmooth part is convex. Surprisingly, unlike the smooth case, our knowledge of this fundamental problem…

Optimization and Control · Mathematics 2016-05-24 Sashank J. Reddi , Suvrit Sra , Barnabas Poczos , Alex Smola

We study the global linear convergence of policy gradient (PG) methods for finite-horizon continuous-time exploratory linear-quadratic control (LQC) problems. The setting includes stochastic LQC problems with indefinite costs and allows…

Optimization and Control · Mathematics 2024-03-05 Michael Giegrich , Christoph Reisinger , Yufei Zhang

We obtain a probabilistic solution to linear-quadratic optimal control problems with state constraints. Given a closed set $\mathcal{D}\subseteq [0,T]\times\mathbb{R}^d$, a diffusion $X$ in $\mathbb{R}^d$ must be linearly controlled in…

Optimization and Control · Mathematics 2026-03-06 Tiziano De Angelis , Erik Ekström

Recent advancements in diffusion models have been effective in learning data priors for solving inverse problems. They leverage diffusion sampling steps for inducing a data prior while using a measurement guidance gradient at each step to…

Machine Learning · Computer Science 2025-04-02 Rayhan Zirvi , Bahareh Tolooshams , Anima Anandkumar

We revisit the finite time analysis of policy gradient methods in the one of the simplest settings: finite state and action MDPs with a policy class consisting of all stochastic policies and with exact gradient evaluations. There has been…

Machine Learning · Computer Science 2021-12-14 Jalaj Bhandari , Daniel Russo