中文
相关论文

相关论文: End-to-End Learning Framework for Solving Non-Mark…

200 篇论文

This paper addresses the problem of safety-critical control for non-affine control systems. It has been shown that optimizing quadratic costs subject to state and control constraints can be sub-optimally reduced to a sequence of quadratic…

系统与控制 · 电气工程与系统科学 2024-02-15 Wei Xiao , Ross Allen , Daniela Rus

Data-driven decision-making processes increasingly utilize end-to-end learnable deep neural networks to render final decisions. Sometimes, the output of the forward functions in certain layers is determined by the solutions to mathematical…

机器学习 · 计算机科学 2024-12-31 Jianming Pan , Zeqi Ye , Xiao Yang , Xu Yang , Weiqing Liu , Lewen Wang , Jiang Bian

We present a control framework for stochastic compartmental models in epidemiology. In this framework, rather than directly controlling the stochastic system, we perform optimal control of an associated Fokker-Planck equation, with the goal…

最优化与控制 · 数学 2026-04-06 Christian Parkinson , Souvik Roy

This paper presents an efficient numerical technique for solving multi-dimensional fractional optimal control problems using fractional-order generalized Bernoulli wavelets. The numerical results obtained by this method have been compared…

最优化与控制 · 数学 2023-10-13 Akanksha Singh , S. Saha Ray

Flying robots such as the quadrotor could provide an efficient approach for medical treatment or sensor placing of wild animals. In these applications, continuously targeting the moving animal is a crucial requirement. Due to the…

机器人学 · 计算机科学 2024-04-16 Ziying Lin , Wei Dong , Sensen Liu , Xinjun Sheng , Xiangyang Zhu

We consider the problem of discounted optimal state-feedback regulation for general unknown deterministic discrete-time systems. It is well known that open-loop instability of systems, non-quadratic cost functions and complex nonlinear…

系统与控制 · 电气工程与系统科学 2020-03-31 Alexandros Tanzanakis , John Lygeros

We study the linear quadratic Gaussian (LQG) control problem, in which the controller's observation of the system state is such that a desired cost is unattainable. To achieve the desired LQG cost, we introduce a communication link from the…

最优化与控制 · 数学 2021-09-28 Oron Sabag , Peida Tian , Victoria Kostina , Babak Hassibi

Over the last few years, sampling-based stochastic optimal control (SOC) frameworks have shown impressive performances in reinforcement learning (RL) with applications in robotics. However, such approaches require a large amount of samples…

系统与控制 · 计算机科学 2014-12-10 Yunpeng Pan , Evangelos A. Theodorou , Michail Kontitsis

In this paper, we investigate a continuous-time linear quadratic control problem for systems with unknown matrices, where only input-output data are available. We propose an output-feedback learning framework based on a canonical nonminimal…

最优化与控制 · 数学 2026-05-19 Weijian Li , Bowen Yi , Panos J. Antsaklis , Hai Lin

Test-time compute has emerged as a powerful paradigm for improving the performance of large language models (LLMs), where generating multiple outputs or refining individual chains can significantly boost answer accuracy. However, existing…

机器学习 · 计算机科学 2025-09-26 Sheng Liu , Tianlang Chen , Pan Lu , Haotian Ye , Yizheng Chen , Lei Xing , James Zou

In this paper, the reinforcement learning (RL)-based optimal control problem is studied for multiplicative-noise systems, where input delay is involved and partial system dynamics is unknown. To solve a variant of Riccati-ZXL equations,…

最优化与控制 · 数学 2023-01-10 Hongxia Wang , Fuyu Zhao , Zhaorong Zhang , Juanjuan Xu , Xun Li

Gradient-based methods have been widely used for system design and optimization in diverse application domains. Recently, there has been a renewed interest in studying theoretical properties of these methods in the context of control and…

最优化与控制 · 数学 2022-10-11 Bin Hu , Kaiqing Zhang , Na Li , Mehran Mesbahi , Maryam Fazel , Tamer Başar

Existing approaches to diffusion-based inverse problem solvers frame the signal recovery task as a probabilistic sampling episode, where the solution is drawn from the desired posterior distribution. This framework suffers from several…

机器学习 · 计算机科学 2024-12-24 Henry Li , Marcus Pereira

Reinforcement learning (RL) has seen significant research and application results but often requires large amounts of training data. This paper proposes two data-efficient off-policy RL methods that use parametrized Q-learning. In these…

系统与控制 · 电气工程与系统科学 2025-04-09 J. S. van Hulst , W. P. M. H. Heemels , D. J. Antunes

We consider the Linear-Quadratic-Regulator (LQR) problem in terms of optimizing a real-valued matrix function over the set of feedback gains. Such a setup facilitates examining the implications of a natural initial-state independent…

系统与控制 · 电气工程与系统科学 2019-07-31 Jingjing Bu , Afshin Mesbahi , Maryam Fazel , Mehran Mesbahi

The fundamental lemma by Jan C. Willems and co-authors enables the representation of all input-output trajectories of a linear time-invariant system by measured input-output data. This result has proven to be pivotal for data-driven…

系统与控制 · 电气工程与系统科学 2024-11-06 Guanru Pan , Ruchuan Ou , Timm Faulwasser

We study the problem of designing a state feedback linear quadratic Gaussian (LQG) controller for a system in which the system matrices as well as the process noise covariance are unknown. We do a rigorous comparison between two approaches.…

系统与控制 · 电气工程与系统科学 2025-11-13 Mingxiang Liu , Damián Marelli , Minyue Fu , Qianqian Cai

This article explores the discrete-time stochastic optimal LQR control with delay and quadratic constraints. The inclusion of delay, compared to delay-free optimal LQR control with quadratic constraints, significantly increases the…

最优化与控制 · 数学 2024-11-19 Dawei Liu , Juanjuan Xu , huanshui Zhang

Fuzzy logic based PID controllers have been studied in this paper, considering several combinations of hybrid controllers by grouping the proportional, integral and derivative actions with fuzzy inferencing in different forms. Fractional…

最优化与控制 · 数学 2013-06-18 Saptarshi Das , Indranil Pan , Shantanu Das

The Linear Quadratic Regulator (LQR) framework considers the problem of regulating a linear dynamical system perturbed by environmental noise. We compute the policy regret between three distinct control policies: i) the optimal online…

最优化与控制 · 数学 2020-02-10 Gautam Goel , Babak Hassibi
‹ 上一页 1 8 9 10 下一页 ›