中文
相关论文

相关论文: Analyzing the Impact of Computation in Adaptive Dy…

200 篇论文

This paper studies a finite-horizon Markov decision problem with information-theoretic constraints, where the goal is to minimize directed information from the controlled source process to the control process, subject to stage-wise cost…

系统与控制 · 电气工程与系统科学 2025-09-04 Zixuan He , Charalambos D. Charalambous , Photios A. Stavrou

The goal of this paper is to develop data-driven control design and evaluation strategies based on linear matrix inequalities (LMIs) and dynamic programming. We consider deterministic discrete-time LTI systems, where the system model is…

最优化与控制 · 数学 2021-06-17 Donghwan Lee , Do Wan Kim

In the context of autonomous driving, the iterative linear quadratic regulator (iLQR) is known to be an efficient approach to deal with the nonlinear vehicle model in motion planning problems. Particularly, the constrained iLQR algorithm…

机器人学 · 计算机科学 2022-07-28 Jun Ma , Zilong Cheng , Xiaoxue Zhang , Masayoshi Tomizuka , Tong Heng Lee

Autonomous robot navigation systems often rely on hierarchical planning, where global planners compute collision-free paths without considering dynamics, and local planners enforce dynamics constraints to produce executable commands. This…

机器人学 · 计算机科学 2025-10-14 Yuanjie Lu , Mingyang Mao , Tong Xu , Linji Wang , Xiaomin Lin , Xuesu Xiao

This paper addresses the abstract dynamic programming (DP) in the online scenario, where the abstract DP mapping is time-varying, instead of static. In this case, optimal costs and policies at different time instants are not the same in…

最优化与控制 · 数学 2021-07-06 Xiuxian Li , Lihua Xie

In order to autonomously learn to control unknown systems optimally w.r.t. an objective function, Adaptive Dynamic Programming (ADP) is well-suited to adapt controllers based on experience from interaction with the system. In recent years,…

系统与控制 · 电气工程与系统科学 2020-02-18 Florian Köpf , Simon Ramsteiner , Michael Flad , Sören Hohmann

The stochastic logistic model with regime switching is an important model in the ecosystem. While analytic solution to this model is positive, current numerical methods are unable to preserve such boundaries in the approximation. So,…

数值分析 · 数学 2021-06-08 Xiaoyue Li , Hongfu Yang

We study the strong approximation of stochastic differential equations with discontinuous drift coefficients and (possibly) degenerate diffusion coefficients. To account for the discontinuity of the drift coefficient we construct an…

数值分析 · 数学 2019-04-25 Andreas Neuenkirch , Michaela Szölgyenyi , Lukasz Szpruch

This paper is concerned with the adaptive numerical treatment of stochastic partial differential equations. Our method of choice is Rothe's method. We use the implicit Euler scheme for the time discretization. Consequently, in each step, an…

In this paper we continue our work on adaptive timestep control for weakly non- stationary problems. The core of the method is a space-time splitting of adjoint error representations for target functionals due to S\"uli and Hartmann. The…

数值分析 · 数学 2014-06-19 Christina Steiner , Siegfried Müller , Sebastian Noelle

We present a receding-horizon optimal control for nonlinear continuous-time systems subject to state constraints. The cost is a quadratic finite-horizon integral. The key enabling technique is a new constrained approximate dynamic…

系统与控制 · 电气工程与系统科学 2026-04-03 Ricardo Gutierrez , Jesse B. Hoagg

This paper addresses the model-free nonlinear optimal problem with generalized cost functional, and a data-based reinforcement learning technique is developed. It is known that the nonlinear optimal control problem relies on the solution of…

系统与控制 · 计算机科学 2013-11-20 Biao Luo , Huai-Ning Wu , Tingwen Huang , Derong Liu

In this paper, we develop a Topological Approximate Dynamic Programming (TADP) method for planningin stochastic systems modeled as Markov Decision Processesto maximize the probability of satisfying high-level systemspecifications expressed…

最优化与控制 · 数学 2020-08-04 Lening Li , Jie Fu

Many applications -- including power systems, robotics, and economics -- involve a dynamical system interacting with a stochastic and hard-to-model environment. We adopt a reinforcement learning approach to control such systems.…

最优化与控制 · 数学 2025-08-26 Abed AlRahman Al Makdah , Oliver Kosut , Lalitha Sankar , Shaofeng Zou

Background. It is assumed that the introduction of stochastic in mathematical model makes it more adequate. But there is virtually no methods of coordinated (depended on structure of the system) stochastic introduction into deterministic…

符号计算 · 计算机科学 2015-03-26 E. G. Eferina , A. V. Korolkova , M. N. Gevorkyan , D. S. Kulyabov , L. A. Sevastyanov

We describe an adaptive importance sampling algorithm for rare events that is based on a dual stochastic control formulation of a path sampling problem. Specifically, we focus on path functionals that have the form of cumulate generating…

动力系统 · 数学 2019-01-30 Omar Kebiri , Lara Neureither , Carsten Hartmann

This work focuses on numerical solutions of optimal control problems. A time discretization error representation is derived for the approximation of the associated value function. It concerns Symplectic Euler solutions of the Hamiltonian…

最优化与控制 · 数学 2016-02-23 Jesper Karlsson , Stig Larsson , Mattias Sandberg , Anders Szepessy , Raùl Tempone

We consider online statistical inference of constrained stochastic nonlinear optimization problems. We apply the Stochastic Sequential Quadratic Programming (StoSQP) method to solve these problems, which can be regarded as applying…

最优化与控制 · 数学 2025-02-19 Sen Na , Michael W. Mahoney

We present an approach for approximately solving discrete-time stochastic optimal-control problems by combining direct trajectory optimization, deterministic sampling, and policy optimization. Our feedback motion-planning algorithm uses a…

机器人学 · 计算机科学 2023-01-12 Taylor A. Howell , Chunjiang Fu , Zachary Manchester

Convex quadratic programs (QPs) constitute a fundamental computational primitive across diverse domains including financial optimization, control systems, and machine learning. The alternating direction method of multipliers (ADMM) has…

最优化与控制 · 数学 2025-05-15 Xi Gao , Jinxin Xiong , Linxin Yang , Akang Wang , Weiwei Xu , Jiang Xue