中文
相关论文

相关论文: Dynamical Low-Rank Approximation Strategies for No…

200 篇论文

This paper offers a unified perspective on different approaches to the solution of optimal control problems through the lens of constrained sequential quadratic programming. In particular, it allows us to find the relationships between…

最优化与控制 · 数学 2025-10-07 Abhijeet , Suman Chakravorty

Deep Reinforcement Learning (DRL) has become a popular method for solving control problems in power systems. Conventional DRL encourages the agent to explore various policies encoded in a neural network (NN) with the goal of maximizing the…

系统与控制 · 电气工程与系统科学 2024-10-28 Tong Wu , Anna Scaglione , Daniel Arnold

We consider the problem of discounted optimal state-feedback regulation for general unknown deterministic discrete-time systems. It is well known that open-loop instability of systems, non-quadratic cost functions and complex nonlinear…

系统与控制 · 电气工程与系统科学 2020-03-31 Alexandros Tanzanakis , John Lygeros

We propose a novel framework for learning a low-dimensional representation of data based on nonlinear dynamical systems, which we call dynamical dimension reduction (DDR). In the DDR model, each point is evolved via a nonlinear flow towards…

机器学习 · 统计学 2022-04-19 Ryeongkyung Yoon , Braxton Osting

Even for known nonlinear dynamical systems, feedback controller synthesis is a difficult problem that often requires leveraging the particular structure of the dynamics to induce a stable closed-loop system. For general nonlinear models,…

系统与控制 · 电气工程与系统科学 2023-06-27 Spencer M. Richards , Jean-Jacques Slotine , Navid Azizan , Marco Pavone

This paper is concerned with a linear-quadratic (LQ, for short) optimal control problem for backward stochastic differential equations (BSDEs, for short), where the coefficients of the backward control system and the weighting matrices in…

最优化与控制 · 数学 2021-05-14 Jingrui Sun , Hanxiao Wang

Alignment is vital for safely deploying large language models (LLMs). Existing techniques are either reward-based (training a reward model on preference pairs and optimizing with reinforcement learning) or reward-free (directly fine-tuning…

计算与语言 · 计算机科学 2026-03-03 Ruoxi Cheng , Haoxuan Ma , Weixin Wang , Ranjie Duan , Jiexi Liu , Xiaoshuang Jia , Simeng Qin , Xiaochun Cao , Yang Liu , Xiaojun Jia

This paper studies the adaptive optimal control problem for a class of linear time-delay systems described by delay differential equations (DDEs). A crucial strategy is to take advantage of recent developments in reinforcement learning and…

系统与控制 · 电气工程与系统科学 2022-10-04 Leilei Cui , Bo Pang , Zhong-Ping Jiang

We study the time-inconsistent linear quadratic optimal control problem for forward-backward stochastic differential equations with potentially indefinite cost weighting matrices for both the state and the control variables. Our research…

最优化与控制 · 数学 2023-12-15 Qi Lü , Bowen Ma

The convergence of policy gradient algorithms hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of linear quadratic control.…

最优化与控制 · 数学 2023-11-02 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

Recently it has been found that for a stochastic linear-quadratic optimal control problem (LQ problem, for short) in a finite horizon, open-loop solvability is strictly weaker than closed-loop solvability which is equivalent to the regular…

最优化与控制 · 数学 2018-06-15 Jingrui Sun , Hanxiao Wang , Jiongmin Yong

We discuss the feedback control problem for a two-dimensional two-phase Stefan problem. In our approach, we use a sharp interface representation in combination with mesh-movement to track the interface position. To attain a feedback…

数值分析 · 数学 2022-12-22 Björn Baran , Peter Benner , Jens Saak

This paper analyzes a special instance of nonsymmetric algebraic matrix Riccati equations arising from transport theory. Traditional approaches for finding the minimal nonnegative solution of the matrix Riccati equations are based on the…

数值分析 · 数学 2011-09-26 Chun-Yueh Chiang , Matthew M. Lin

This paper develops a data-based approach to the closed-loop output feedback control of nonlinear dynamical systems with a partial nonlinear observation model. We propose an information state based approach to rigorously transform the…

机器人学 · 计算机科学 2023-10-06 Raman Goyal , Ran Wang , Mohamed Naveed Gul Mohamed , Aayushman Sharma , Suman Chakravorty

We propose a two-phase risk-averse architecture for controlling stochastic nonlinear robotic systems. We present Risk-Averse Nonlinear Steering RRT* (RANS-RRT*) as an RRT* variant that incorporates nonlinear dynamics by solving a nonlinear…

机器人学 · 计算机科学 2021-09-07 Sleiman Safaoui , Benjamin J. Gravell , Venkatraman Renganathan , Tyler H. Summers

Quantifying uncertainties in hyperbolic equations is a source of several challenges. First, the solution forms shocks leading to oscillatory behaviour in the numerical approximation of the solution. Second, the number of unknowns required…

数值分析 · 数学 2021-05-11 Jonas Kusch , Gianluca Ceruti , Lukas Einkemmer , Martin Frank

We are concerned with efficient numerical methods for stochastic continuous-time algebraic Riccati equations (SCARE). Such equations frequently arise from the state-dependent Riccati equation approach which is perhaps the only systematic…

最优化与控制 · 数学 2024-01-23 Tsung-Ming Huang , Yueh-Cheng Kuo , Ren-Cang Li , Wen-Wei Lin

Dynamic Traffic Assignment (DTA) provides an approach to determine the optimal path and/or departure time based on the transportation network characteristics and user behavior (e.g., selfish or social). In the literature, most of the…

系统与控制 · 计算机科学 2017-08-15 Tarikul Islam , Hai L. Vu , Manoj Panda , Nam Hoang , Dong Ngoduy

Computing effective eigenvalues for neutron transport often requires a fine numerical resolution. The main challenge of such computations is the high memory effort of classical solvers, which limits the accuracy of chosen discretizations.…

数值分析 · 数学 2022-09-28 Jonas Kusch , Benjamin Whewell , Ryan McClarren , Martin Frank

This is a draft paper originally posted on Arxiv as a documentation of a plenary lecture at CDC2023. The core material has been accepted for publication at L4DC 2024. Certainty equivalence adaptive controllers are analysed using a…

最优化与控制 · 数学 2024-06-04 Anders Rantzer