English
Related papers

Related papers: How are policy gradient methods affected by the li…

200 papers

Prescribed-time algorithms based on time-varying gains may have remarkable properties, such as regulation in a user-prescribed finite time that is the same for every nonzero initial condition and that holds even under matched disturbances.…

Systems and Control · Electrical Eng. & Systems 2023-12-18 Rodrigo Aldana-López , Richard Seeber , Hernan Haimovich , David Gómez-Gutiérrez

The problem of synthesizing stochastic explicit model predictive control policies is known to be quickly intractable even for systems of modest complexity when using classical control-theoretic methods. To address this challenge, we present…

Machine Learning · Computer Science 2022-05-24 Ján Drgoňa , Sayak Mukherjee , Aaron Tuor , Mahantesh Halappanavar , Draguna Vrabie

In this work we provide a computationally tractable procedure for designing affine control policies, applied to constrained, discrete-time, partially observable, linear systems subject to set bounded disturbances, stochastic noise and…

Optimization and Control · Mathematics 2018-11-27 Georgios Kotsalis , Guanghui Lan

In view of solving convex optimization problems with noisy gradient input, we analyze the asymptotic behavior of gradient-like flows under stochastic disturbances. Specifically, we focus on the widely studied class of mirror descent schemes…

Optimization and Control · Mathematics 2017-09-21 Panayotis Mertikopoulos , Mathias Staudigl

Nonlinear systems are often subject to random influences. Sometimes the noise enters the system through physical boundaries and this leads to stochastic dynamic boundary conditions. A dynamic, as opposed to static, boundary condition…

Dynamical Systems · Mathematics 2007-05-23 Desheng Yang , Jinqiao Duan

Most biological systems are formed by component parts that to some degree are inter-related. Groups of parts that are more associated among themselves and are relatively autonomous from others are called modules. One of the consequences of…

Populations and Evolution · Quantitative Biology 2013-08-12 Gabriel Marroig , Diogo Melo , Guilherme Garcia

We consider a 2-dimensional stochastic differential equation in polar coordinates depending on several parameters. We show that if these parameters belong to a specific regime then the deterministic system explodes in finite time, but the…

Dynamical Systems · Mathematics 2022-06-17 Matti Leimbach , Jonathan C. Mattingly , Michael Scheutzow

Increasing effort is put into the development of methods for learning mechanistic models from data. This task entails not only the accurate estimation of parameters but also a suitable model structure. Recent work on the discovery of…

Machine Learning · Computer Science 2024-07-01 Justin N. Kreikemeyer , Philipp Andelfinger , Adelinde M. Uhrmacher

Linear functions of many independent random variables lead to classical noises (white, Poisson, and their combinations) in the scaling limit. Some singular stochastic flows and some models of oriented percolation involve very nonlinear…

Probability · Mathematics 2007-05-23 Boris Tsirelson

We present a full stochastic description of the pair approximation scheme to study binary-state dynamics on heterogeneous networks. Within this general approach, we obtain a set of equations for the dynamical correlations, fluctuations and…

Physics and Society · Physics 2018-11-05 A. F. Peralta , A. Carro , M. San Miguel , R. Toral

This note is addressed to giving a short introduction to control theory of stochastic systems, governed by stochastic differential equations in both finite and infinite dimensions. We will mainly explain the new phenomenon and difficulties…

Optimization and Control · Mathematics 2016-12-09 Qi Lu , Xu Zhang

Policy gradient methods are widely used in reinforcement learning. Yet, the nonconvexity of policy optimization poses significant challenges in understanding the global convergence of policy gradient methods. For a class of finite-horizon…

Optimization and Control · Mathematics 2026-03-10 Xin Chen , Yifan Hu , Minda Zhao

In this work, we consider a multi-population system where the dynamics of each agent evolve according to a system of stochastic differential equations in a general functional setup, determined by the global state of the system. Each agent…

Probability · Mathematics 2025-07-24 Giuseppe D'Onofrio , Anderson Melchor Hernandez

The policy gradient theorem describes the gradient of the expected discounted return with respect to an agent's policy parameters. However, most policy gradient methods drop the discount factor from the state distribution and therefore do…

Machine Learning · Computer Science 2020-03-02 Chris Nota , Philip S. Thomas

This work is concerned with existence of weak solutions to discon- tinuous stochastic differential equations driven by multiplicative Gaus- sian noise and sliding mode control dynamics generated by stochastic differential equations with…

Optimization and Control · Mathematics 2015-04-27 Viorel Barbu , Stefano Bonaccorsi , Luciano Tubaro

This paper examines learning the optimal filtering policy, known as the Kalman gain, for a linear system with unknown noise covariance matrices using noisy output data. The learning problem is formulated as a stochastic policy optimization…

Systems and Control · Electrical Eng. & Systems 2023-10-27 Shahriar Talebi , Amirhossein Taghvaei , Mehran Mesbahi

To sample from an unconditionally trained Denoising Diffusion Probabilistic Model (DDPM), classifier guidance adds conditional information during sampling, but the gradients from classifiers, especially those not trained on noisy images,…

Machine Learning · Computer Science 2024-06-26 Philipp Vaeth , Alexander M. Fruehwald , Benjamin Paassen , Magda Gregorova

In this paper, we present a methodology to deploy the deterministic policy gradient method, using actor-critic techniques, when the optimal policy is approximated using a parametric optimization problem, where safety is enforced via hard…

Systems and Control · Electrical Eng. & Systems 2024-09-23 Sebastien Gros , Mario Zanon

We find that the performance of state-of-the-art models on Natural Language Inference (NLI) and Reading Comprehension (RC) analysis/stress sets can be highly unstable. This raises three questions: (1) How will the instability affect the…

Computation and Language · Computer Science 2020-11-17 Xiang Zhou , Yixin Nie , Hao Tan , Mohit Bansal

Stochastic inverse problems considered in this article consist of estimating the probability distributions of intrinsically random inputs of computer models. These estimations are based on observable outputs affected by model noise, and…

Statistics Theory · Mathematics 2025-03-17 Nicolas Bousquet , Mélanie Blazère , Thomas Cerbelaud