中文
相关论文

相关论文: Memory-Limited Partially Observable Stochastic Con…

200 篇论文

In model predictive control (MPC), the choice of cost-weighting matrices and designing the Hessian matrix directly affects the trade-off between rapid state regulation and minimizing the control effort. However, traditional MPC in quadratic…

系统与控制 · 电气工程与系统科学 2026-02-13 Komeil Nosrati , Juri Belikov , Aleksei Tepljakov , Eduard Petlenkov

Partially Observable Markov Decision Processes (POMDPs) are a fundamental framework for decision-making under uncertainty and partial observability. Since in general optimal policies may require infinite memory, they are hard to implement…

人工智能 · 计算机科学 2026-04-30 Muqsit Azeem , Debraj Chakraborty , Sudeep Kanav , Jan Kretinsky

In this paper, we propose a combined Magnitude Saturated Adaptive Control (MSAC)-Model Predictive Control (MPC) approach to linear quadratic tracking optimal control problems with parametric uncertainties and input saturation. The proposed…

最优化与控制 · 数学 2023-03-14 Sunbochen Tang , Anuradha M. Annaswamy

Reinforcement learning in partially observed Markov decision processes (POMDPs) faces two challenges. (i) It often takes the full history to predict the future, which induces a sample complexity that scales exponentially with the horizon.…

机器学习 · 计算机科学 2024-04-02 Lingxiao Wang , Qi Cai , Zhuoran Yang , Zhaoran Wang

This paper shows that the optimal policy and value functions of a Markov Decision Process (MDP), either discounted or not, can be captured by a finite-horizon undiscounted Optimal Control Problem (OCP), even if based on an inexact model.…

系统与控制 · 电气工程与系统科学 2023-02-08 Arash Bahari Kordabad , Mario Zanon , Sebastien Gros

Mean-Field Control (MFC) has recently been proven to be a scalable tool to approximately solve large-scale multi-agent reinforcement learning (MARL) problems. However, these studies are typically limited to unconstrained cumulative reward…

机器学习 · 计算机科学 2024-09-11 Washim Uddin Mondal , Vaneet Aggarwal , Satish V. Ukkusuri

The partially observable constrained optimization problems (POCOPs) impede data-driven optimization techniques since an infeasible solution of POCOPs can provide little information about the objective as well as the constraints. We endeavor…

机器学习 · 计算机科学 2023-12-27 Shengbo Wang , Ke Li

Linear-Quadratic-Gaussian (LQG) control is concerned with the design of an optimal controller and estimator for linear Gaussian systems with imperfect state information. Standard LQG assumes the set of sensor measurements, to be fed to the…

最优化与控制 · 数学 2020-05-18 Vasileios Tzoumas , Luca Carlone , George J. Pappas , Ali Jadbabaie

Partially observable Markov decision processes (POMDPs) are a natural model for planning problems where effects of actions are nondeterministic and the state of the world is not completely observable. It is difficult to solve POMDPs…

人工智能 · 计算机科学 2009-09-25 N. L. Zhang , W. Liu

This paper studies online solutions for regret-optimal control in partially observable systems over an infinite-horizon. Regret-optimal control aims to minimize the difference in LQR cost between causal and non-causal controllers while…

系统与控制 · 电气工程与系统科学 2023-11-15 Joudi Hajar , Oron Sabag , Babak Hassibi

We study methods for solving stochastic control problems of systems of forward-backward mean-field equations with delay, in finite or infinite horizon. Necessary and sufficient maximum principles under partial information are given. The…

最优化与控制 · 数学 2016-10-31 Nacira Agram , Elin Engen Rose

We revisit closed-loop performance guarantees for Model Predictive Control in the deterministic and stochastic cases, which extend to novel performance results applicable to receding horizon control of Partially Observable Markov Decision…

最优化与控制 · 数学 2020-05-01 Martin A. Sehr , Robert R. Bitmead

The article poses a general model for optimal control subject to information constraints, motivated in part by recent work of Sims and others on information-constrained decision-making by economic agents. In the average-cost optimal control…

最优化与控制 · 数学 2016-02-24 Ehsan Shafieepoorfard , Maxim Raginsky , Sean P. Meyn

Mean-Field Control (MFC) is a powerful tool to solve Multi-Agent Reinforcement Learning (MARL) problems. Recent studies have shown that MFC can well-approximate MARL when the population size is large and the agents are exchangeable.…

机器学习 · 计算机科学 2022-06-02 Washim Uddin Mondal , Vaneet Aggarwal , Satish V. Ukkusuri

This paper is concerned with a class of linear-quadratic stochastic large-population problems with partial information, where the individual agent only has access to a noisy observation process related to the state. The dynamics of each…

最优化与控制 · 数学 2024-08-20 Min Li , Na Li , Zhen Wu

In this work, we study a class of mean-field linear quadratic Gaussian (LQG) problems. Under suitable conditions, explicit solutions of the distribution-dependent optimal control problems are obtained. Riccati systems are derived by…

概率论 · 数学 2020-08-28 Yun Li , Qingshuo Song , Fuke Wu , George Yin

Partially Observable Monte-Carlo Planning (POMCP) is a powerful online algorithm able to generate approximate policies for large Partially Observable Markov Decision Processes. The online nature of this method supports scalability by…

人工智能 · 计算机科学 2021-04-29 Giulio Mazzi , Alberto Castellini , Alessandro Farinelli

In this paper, we present a multilevel Monte Carlo (MLMC) version of the Stochastic Gradient (SG) method for optimization under uncertainty, in order to tackle Optimal Control Problems (OCP) where the constraints are described in the form…

最优化与控制 · 数学 2019-12-30 Matthieu Martin , Fabio Nobile , Panagiotis Tsilifis

Based on our recent research on neural heuristic quantization systems, we propose an emulation problem consistent with the neuromimetic paradigm. This optimal quantization problem can be solved with model predictive control (MPC) by…

系统与控制 · 电气工程与系统科学 2023-05-08 Zexin Sun , John Baillieul

The current study addresses the control problems posed by a semilinear neutral integro-differential equation with memory. The primary objectives of this study are to investigate the existence of a mild solution and approximate…

最优化与控制 · 数学 2025-01-28 Sumit Arora , Akambadath Nandakumaran