English
Related papers

Related papers: Maximum Principle Based Algorithms for Deep Learni…

200 papers

Training the deep convolutional neural network for computer vision problems is slow and inefficient, especially when it is large and distributed across multiple devices. The inefficiency is caused by the backpropagation algorithm's forward…

Machine Learning · Computer Science 2022-01-20 An Xu , Zhouyuan Huo , Heng Huang

In this paper, we introduce proximal gradient temporal difference learning, which provides a principled way of designing and analyzing true stochastic gradient temporal difference learning algorithms. We show how gradient TD (GTD)…

Machine Learning · Computer Science 2020-06-09 Bo Liu , Ian Gemp , Mohammad Ghavamzadeh , Ji Liu , Sridhar Mahadevan , Marek Petrik

We study a Q learning algorithm for continuous time stochastic control problems. The proposed algorithm uses the sampled state process by discretizing the state and control action spaces under piece-wise constant control processes. We show…

Optimization and Control · Mathematics 2023-03-10 Erhan Bayraktar , Ali Devran Kara

We propose a Safe Pontryagin Differentiable Programming (Safe PDP) methodology, which establishes a theoretical and algorithmic framework to solve a broad class of safety-critical learning and control tasks -- problems that require the…

Machine Learning · Computer Science 2021-10-27 Wanxin Jin , Shaoshuai Mou , George J. Pappas

This paper provides a review and commentary on the past, present, and future of numerical optimization algorithms in the context of machine learning applications. Through case studies on text classification and the training of deep neural…

Machine Learning · Statistics 2018-02-12 Léon Bottou , Frank E. Curtis , Jorge Nocedal

Lagrangian systems represent a wide range of robotic systems, including manipulators, wheeled and legged robots, and quadrotors. Inverse dynamics control and feedforward linearization techniques are typically used to convert the complex…

Robotics · Computer Science 2018-09-13 Mohamed K. Helwa , Adam Heins , Angela P. Schoellig

Infinite-time nonlinear optimal regulation control is widely utilized in aerospace engineering as a systematic method for synthesizing stable controllers. However, conventional methods often rely on linearization hypothesis, while recent…

Systems and Control · Electrical Eng. & Systems 2025-06-13 Han Wang , Di Wu , Lin Cheng , Shengping Gong , Xu Huang

Enforcing state and input constraints during reinforcement learning (RL) in continuous state spaces is an open but crucial problem which remains a roadblock to using RL in safety-critical applications. This paper leverages invariant sets to…

Systems and Control · Electrical Eng. & Systems 2019-06-28 Ankush Chakrabarty , Rien Quirynen , Claus Danielson , Weinan Gao

In this article we derive a strong version of the Pontryagin Maximum Principle for general nonlinear optimal control problems on time scales in finite dimension. The final time can be fixed or not, and in the case of general boundary…

Optimization and Control · Mathematics 2013-02-15 Loïc Bourdin , Emmanuel Trélat

We obtain the dynamic programming equations and optimality conditions akin to Pontryagin's extremum principle for certain mathematical models of hybrid control systems.

Optimization and Control · Mathematics 2007-05-23 S. A. Belbas

We establish a Pontryagin maximum principle for discrete time optimal control problems under the following three types of constraints: a) constraints on the states pointwise in time, b) constraints on the control actions pointwise in time,…

Optimization and Control · Mathematics 2019-05-27 Pradyumna Paruchuri , Debasish Chatterjee

Stochastic differential equations can describe a wide range of dynamical systems, and obtaining the governing equations of these systems is the premise of studying the nonlinear dynamic behavior of the system. Neural networks are currently…

Dynamical Systems · Mathematics 2023-04-25 Xiao-Kai An , Lin Du , Zi-Chen Deng , Yu-jia Zhang

We analyze the convergence rate of various momentum-based optimization algorithms from a dynamical systems point of view. Our analysis exploits fundamental topological properties, such as the continuous dependence of iterates on their…

Optimization and Control · Mathematics 2021-04-13 Michael Muehlebach , Michael I. Jordan

Even nowadays, where Deep Learning (DL) has achieved state-of-the-art performance in a wide range of research domains, accelerating training and building robust DL models remains a challenging task. To this end, generations of researchers…

Machine Learning · Computer Science 2024-08-22 Manos Kirtas , Nikolaos Passalis , Anastasios Tefas

This paper develops a primal-dual dynamical system where the coefficients are designed in closed-loop way for solving a convex optimization problem with linear equality constraints. We first introduce a ``second-order primal" +…

Optimization and Control · Mathematics 2026-03-03 Huan Zhang , Xiangkai Sun , Shengjie Li , Kok Lay Teo

With the widespread adoption of machine learning systems, the need to curtail their behavior has become increasingly apparent. This is evidenced by recent advancements towards developing models that satisfy robustness, safety, and fairness…

Machine Learning · Computer Science 2024-03-19 Juan Elenter , Luiz F. O. Chamon , Alejandro Ribeiro

In this article, we propose a novel pessimism-based Bayesian learning method for optimal dynamic treatment regimes in the offline setting. When the coverage condition does not hold, which is common for offline data, the existing solutions…

Machine Learning · Statistics 2023-02-23 Yunzhe Zhou , Zhengling Qi , Chengchun Shi , Lexin Li

We extend the Longstaff-Schwartz algorithm for approximately solving optimal stopping problems on high-dimensional state spaces. We reformulate the optimal stopping problem for Markov processes in discrete time as a generalized statistical…

Probability · Mathematics 2007-05-23 Daniel Egloff

The fundamental theorem of the theory of optimal control, the Pontryagin maximum principle (PMP), is extended to the setting of almost Lie (AL) algebroids, geometrical objects generalizing Lie algebroids. This formulation of the PMP yields,…

Optimization and Control · Mathematics 2013-06-13 Janusz Grabowski , Michal Jozwikowski

In this paper, we consider the linear programming (LP) formulation for deep reinforcement learning. The number of the constraints depends on the size of state and action spaces, which makes the problem intractable in large or continuous…

Optimization and Control · Mathematics 2021-05-21 Yongfeng Li , Mingming Zhao , Weijie Chen , Zaiwen Wen