中文
相关论文

相关论文: How are policy gradient methods affected by the li…

200 篇论文

Gradient-based methods have been widely used for system design and optimization in diverse application domains. Recently, there has been a renewed interest in studying theoretical properties of these methods in the context of control and…

最优化与控制 · 数学 2022-10-11 Bin Hu , Kaiqing Zhang , Na Li , Mehran Mesbahi , Maryam Fazel , Tamer Başar

The performance of model-based control techniques strongly depends on the quality of the employed dynamics model. If strong guarantees are desired, it is therefore common to robustly treat all possible sources of uncertainty, such as model…

系统与控制 · 电气工程与系统科学 2022-05-23 Elena Arcari , Andrea Iannelli , Andrea Carron , Melanie N. Zeilinger

This article is concerned with stability and performance of controlled stochastic processes under receding horizon policies. We carry out a systematic study of methods to guarantee stability under receding horizon policies via appropriate…

系统与控制 · 计算机科学 2017-11-27 Debasish Chatterjee , John Lygeros

Feedback control (based on the quantum continuous measurement) of quantum systems inevitably suffers from estimation delays. In this paper we give a delay-dependent stability criterion for a wide class of nonlinear stochastic systems…

量子物理 · 物理学 2011-12-05 Kenji Kashima , Naoki Yamamoto

Policy gradient methods are reinforcement learning algorithms that adapt a parameterized policy by following a performance gradient estimate. Conventional policy gradient methods use Monte-Carlo techniques to estimate the gradient, which…

机器学习 · 计算机科学 2026-05-01 Mohammad Ghavamzadeh , Yaakov Engel , Michal Valko

In this paper, we characterize the noise of stochastic gradients and analyze the noise-induced dynamics during training deep neural networks by gradient-based optimizers. Specifically, we firstly show that the stochastic gradient noise…

机器学习 · 计算机科学 2021-09-22 Yixin Wu , Rui Luo , Chen Zhang , Jun Wang , Yaodong Yang

Noisy fluctuations are ubiquitous in complex systems. They play a crucial or delicate role in the dynamical evolution of gene regulation, signal transduction, biochemical reactions, among other systems. Therefore, it is essential to…

动力系统 · 数学 2018-11-05 Jinqiao Duan , Hui Wang

We consider the control of semilinear stochastic partial differential equations (SPDEs) via deterministic controls. In the case of multiplicative noise, existence of optimal controls and necessary conditions for optimality are derived. In…

最优化与控制 · 数学 2021-10-28 Wilhelm Stannat , Lukas Wessels

Policy gradient methods hold great potential for solving complex continuous control tasks. Still, their training efficiency can be improved by exploiting structure within the optimization problem. Recent work indicates that supervised…

A recent literature considers causal inference using noisy proxies for unobserved confounding factors. The proxies are divided into two sets that are independent conditional on the confounders. One set of proxies are `negative control…

计量经济学 · 经济学 2021-10-11 Ben Deaner

We study a 1D ring of diffusively coupled logistic maps in the vicinity of an unstable, spatially homogeneous fixed point. The failure of linear controllers due to additive noise is discussed with the aim of clarifying the failure…

chao-dyn · 物理学 2009-10-30 David A. Egolf , Joshua E. S. Socolar

The paper considers a stabilizing stochastic control which can be applied to a variety of unstable and even chaotic maps. Compared to previous methods introducing control by noise, we relax assumptions on the class of maps, as well as…

动力系统 · 数学 2019-02-25 Elena Braverman , Alexandra Rodkina

In policy gradient reinforcement learning, access to a differentiable model enables 1st-order gradient estimation that accelerates learning compared to relying solely on derivative-free 0th-order estimators. However, discontinuous dynamics…

机器学习 · 计算机科学 2026-04-21 Ku Onoda , Paavo Parmas , Manato Yaguchi , Yutaka Matsuo

We analyze the complexity of biased stochastic gradient methods (SGD), where individual updates are corrupted by deterministic, i.e. biased error terms. We derive convergence results for smooth (non-convex) functions and give improved rates…

机器学习 · 计算机科学 2021-05-11 Ahmad Ajalloeian , Sebastian U. Stich

We construct control policies that ensure bounded variance of a noisy marginally stable linear system in closed-loop. It is assumed that the noise sequence is a mutually independent sequence of random vectors, enters the dynamics affinely,…

Policy gradient (PG) methods are the backbone of many reinforcement learning algorithms due to their good performance in policy optimization problems. As a gradient-based approach, PG methods typically rely on knowledge of the system…

系统与控制 · 电气工程与系统科学 2026-04-02 Bowen Song , Andrea Iannelli

Designing feasible control strategies for opinion dynamics in complex social systems has never been an easy task. It requires a control protocol which 1) is not enforced on all individuals in the society, and 2) does not exclusively rely on…

物理与社会 · 物理学 2019-02-27 Wei Su , Xianzhong Chen , Yongguang Yu , Ge Chen

In this study, we introduce a sensitivity analysis methodology for stochastic systems in chemistry, where dynamics are often governed by random processes. Our approach is based on gradient estimation via finite differences, averaging…

定量方法 · 定量生物学 2026-01-12 Erika M. Herrera Machado , Jakob L. Andersen , Rolf Fagerberg , Daniel Merkle

This paper examines stochastic optimal control problems in which the state is perfectly known, but the controller's measure of time is a stochastic process derived from a strictly increasing L\'evy process. We provide dynamic programming…

最优化与控制 · 数学 2014-01-03 Andrew Lamperski , Noah J. Cowan

This paper studies the set of terminal state covariances that are reachable over a finite time horizon from a given initial state covariance for a linear stochastic system with additive noise. For discrete-time systems, a complete…

系统与控制 · 电气工程与系统科学 2025-09-22 Fengjiao Liu , Panagiotis Tsiotras