中文
相关论文

相关论文: Initialization-driven neural generation and traini…

200 篇论文

We develop a framework for the analysis of deep neural networks and neural ODE models that are trained with stochastic gradient algorithms. We do that by identifying the connections between control theory, deep learning and theory of…

概率论 · 数学 2021-03-18 Jean-François Jabir , David Šiška , Łukasz Szpruch

A learning based method for obtaining feedback laws for nonlinear optimal control problems is proposed. The learning problem is posed such that the open loop value function is its optimal solution. This infinite dimensional, function space,…

最优化与控制 · 数学 2022-10-26 Karl Kunisch , Donato Vásquez-Varas , Daniel Walter

In this paper, we aim to solve the high dimensional stochastic optimal control problem from the view of the stochastic maximum principle via deep learning. By introducing the extended Hamiltonian system which is essentially an FBSDE with a…

最优化与控制 · 数学 2021-06-23 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

Most existing neural network-based approaches for solving stochastic optimal control problems using the associated backward dynamic programming principle rely on the ability to simulate the underlying state variables. However, in some…

机器学习 · 统计学 2024-01-30 Christian Yeo

Infinite-time nonlinear optimal regulation control is widely utilized in aerospace engineering as a systematic method for synthesizing stable controllers. However, conventional methods often rely on linearization hypothesis, while recent…

系统与控制 · 电气工程与系统科学 2025-06-13 Han Wang , Di Wu , Lin Cheng , Shengping Gong , Xu Huang

In this note, we study a class of indefinite stochastic McKean-Vlasov linear-quadratic (LQ in short) control problem under the control taking nonnegative values. In contrast to the conventional issue, both the classical dynamic programming…

最优化与控制 · 数学 2023-10-05 Xun Li , Liangquan Zhang

Mean field optimal control problems are a class of optimization problems that arise from optimal control when applied to the many body setting. In the noisy case one has a set of controllable stochastic processes and a cost function that is…

最优化与控制 · 数学 2021-08-11 Pierfrancesco Urbani

We propose a mean-field optimal control problem for the parameter identification of a given pattern. The cost functional is based on the Wasserstein distance between the probability measures of the modeled and the desired patterns. The…

最优化与控制 · 数学 2021-04-08 Martin Burger , Lisa Maria Kreusser , Claudia Totzeck

We propose a machine learning algorithm for solving finite-horizon stochastic control problems based on a deep neural network representation of the optimal policy functions. The algorithm has three features: (1) It can solve…

综合经济学 · 经济学 2024-12-09 Xianhua Peng , Steven Kou , Lekang Zhang

We consider the approximation of some optimal control problems for the Navier-Stokes equation via a Dynamic Programming approach. These control problems arise in many industrial applications and are very challenging from the numerical point…

最优化与控制 · 数学 2022-07-18 Maurizio Falcone , Gerhard Kirsten , Luca Saluzzi

This work proposes an optimal safe controller minimizing an infinite horizon cost functional subject to control barrier functions (CBFs) safety conditions. The constrained optimal control problem is reformulated as a minimization problem of…

系统与控制 · 电气工程与系统科学 2022-02-03 Hassan Almubarak , Evangelos A. Theodorou , Nader Sadegh

This paper presents an inverse optimality method to solve the Hamilton-Jacobi-Bellman equation for a class of nonlinear problems for which the cost is quadratic and the dynamics are affine in the input. The method is inverse optimal because…

最优化与控制 · 数学 2011-10-11 Luis Rodrigues , Didier Henrion , Mehdi Abedinpour Fallah

We study optimal control problems in infinite horizon when the dynamics belong to a specific class of piecewise deterministic Markov processes constrained to star-shaped networks (inspired by traffic models). We adapt the results in [H. M.…

最优化与控制 · 数学 2015-10-06 Dan Goreac , Magdalena Kobylanski , Miguel Martinez

We consider initial value problems of nonlinear dynamical systems, which include physical parameters. A quantity of interest depending on the solution is observed. A discretisation yields the trajectories of the quantity of interest in many…

机器学习 · 计算机科学 2021-01-13 Roland Pulch , Maha Youssef

We study the problem of optimal portfolio selection under stochastic volatility within a continuous time reinforcement learning framework with portfolio constraints. Exploration is modeled through entropy-regularized relaxed controls, where…

数理金融 · 定量金融 2026-04-27 Thai Nguyen , Pertiny Nkuize

Deep learning is formulated as a discrete-time optimal control problem. This allows one to characterize necessary conditions for optimality and develop training algorithms that do not rely on gradients with respect to the trainable…

机器学习 · 计算机科学 2018-06-05 Qianxiao Li , Shuji Hao

We propose two algorithms for the solution of the optimal control of ergodic McKean-Vlasov dynamics. Both algorithms are based on approximations of the theoretical solutions by neural networks, the latter being characterized by their…

最优化与控制 · 数学 2021-03-30 René Carmona , Mathieu Laurière

This paper studies optimal consensus tracking problem of heterogeneous linear multi-agent systems. By introducing tracking error dynamics, the optimal tracking problem is reformulated as finding a Nash-equilibrium solution of a multi-player…

最优化与控制 · 数学 2019-05-21 Jilie Zhang , Zhanshan Wang , Hongwei Zhang

In this paper, we first establish the dynamic programming principle for stochastic optimal control problems defined on compact Riemannian manifolds without boundary. Subsequently, we derive the associated Hamilton-Jacobi-Bellman (HJB)…

最优化与控制 · 数学 2025-07-03 Dingqian Gao , Qi Lü

In this paper, we study a time-inconsistent stochastic optimal control problem with a recursive cost functional by a multi-person hierarchical differential game approach. An equilibrium strategy of this problem is constructed and a…

最优化与控制 · 数学 2016-06-13 Qingmeng Wei , Jiongmin Yong , Zhiyong Yu