English
Related papers

Related papers: Bootstrap Policy Iteration for Stochastic LQ Track…

200 papers

This paper is concerned with a linear quadratic (LQ, for short) optimal control problem for mean-field backward stochastic differential equations (MF-BSDE, for short) driven by a Poisson random martingale measure and a Brownian motion.…

Optimization and Control · Mathematics 2016-11-22 Maoning Tang , Qingxin Meng

We consider the problem of robotic planning under uncertainty in this paper. This problem may be posed as a stochastic optimal control problem, a solution to which is fundamentally intractable owing to the infamous "curse of…

Systems and Control · Electrical Eng. & Systems 2019-09-19 Mohamed Naveed Gul Mohamed , Suman Chakravorty , Dylan A. Shell

A robustly stabilizing optimal control policy in a model-free mixed $\mathcal{H}_2/\mathcal{H}_\infty$-control setting is here put forward for counterbalancing the slow convergence and non-robustness of traditional high-variance policy…

Optimization and Control · Mathematics 2023-04-18 Lekan Molu

This paper studies an optimal control problem for continuous-time stochastic systems subject to reachability objectives specified in a subclass of metric interval temporal logic specifications, a temporal logic with real-time constraints.…

Systems and Control · Computer Science 2015-04-21 Jie Fu , Ufuk Topcu

Traditional approaches to motion modeling for skid-steer robots struggle with capturing nonlinear tire-terrain dynamics, especially during high-speed maneuvers. In this paper, we tackle such nonlinearities by enhancing a dynamic unicycle…

Robotics · Computer Science 2024-11-06 Ananya Trivedi , Sarvesh Prajapati , Anway Shirgaonkar , Mark Zolotas , Taskin Padir

Many robotic systems must follow planned paths yet pause safely and resume when people or objects intervene. We present an output-space method for systems whose tracked output can be feedback-linearized to a double integrator (e.g.,…

Robotics · Computer Science 2025-09-18 Hossein Gholampour , Logan E. Beaver

This paper proposes an agent-based optimistic policy iteration (OPI) scheme for learning stationary optimal stochastic policies in multi-agent Markov Decision Processes (MDPs), in which agents incur a Kullback-Leibler (KL) divergence cost…

Artificial Intelligence · Computer Science 2024-10-22 Khaled Nakhleh , Ceyhun Eksin , Sabit Ekin

We study the problem of devising a closed-loop strategy to control the position of a robot that is tracking a possibly moving target. The robot is capable of obtaining noisy measurements of the target's position. The key idea in active…

Robotics · Computer Science 2016-11-09 Zhonghshun Zhang , Pratap Tokekar

We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem. In particular, we consider the convergence of policy gradient methods in the setting of known and unknown parameters.…

Machine Learning · Computer Science 2021-06-25 Ben Hambly , Renyuan Xu , Huining Yang

We study the problem of Safe Policy Improvement (SPI) under constraints in the offline Reinforcement Learning (RL) setting. We consider the scenario where: (i) we have a dataset collected under a known baseline policy, (ii) multiple reward…

Machine Learning · Computer Science 2021-11-01 Harsh Satija , Philip S. Thomas , Joelle Pineau , Romain Laroche

Noise-induced dynamics of a prototypical bistable system with delayed feedback is studied theoretically and numerically. For small noise and magnitude of the feedback, the problem is reduced to the analysis of the two-state model with…

Statistical Mechanics · Physics 2009-11-07 L. S. Tsimring , A. Pikovsky

We present a unified framework for learning continuous control policies using backpropagation. It supports stochastic control by treating stochasticity in the Bellman equation as a deterministic function of exogenous noise. The product is a…

Machine Learning · Computer Science 2015-11-02 Nicolas Heess , Greg Wayne , David Silver , Timothy Lillicrap , Yuval Tassa , Tom Erez

In this article we consider a stochastic optimal control problem where the dynamics of the state process, $X(t)$, is a controlled stochastic differential equation with jumps, delay and \emph{noisy memory}. The term noisy memory is, to the…

Optimization and Control · Mathematics 2015-08-28 Kristina R. Dahl , Salah-Eldin A. Mohammed , Bernt Øksendal , Elin Røse

We present an approach for approximately solving discrete-time stochastic optimal-control problems by combining direct trajectory optimization, deterministic sampling, and policy optimization. Our feedback motion-planning algorithm uses a…

Robotics · Computer Science 2023-01-12 Taylor A. Howell , Chunjiang Fu , Zachary Manchester

This paper first presents necessary and sufficient conditions for the solvability of discrete time, mean-field, stochastic linear-quadratic optimal control problems. Then, by introducing several sequences of bounded linear operators, the…

Optimization and Control · Mathematics 2016-07-25 Robert. J Elliott , Xun Li , Yuan-Hua Ni

Learning-based methods are powerful in handling complex scenarios. However, it is still challenging to use learning-based methods under uncertain environments while stability, safety, and real-time performance of the system are desired to…

Robotics · Computer Science 2022-03-08 Zhixuan Wu , Rui Yang , Lei Zheng , Hui Cheng

This paper focuses on the linear quadratic control (LQC) design of systems corrupted by both stochastic noise and bounded noise simultaneously. When only of these noises are considered, the LQC strategy leads to stochastic or robust…

Optimization and Control · Mathematics 2025-12-15 Xuehui Ma , Shiliang Zhang , Xiaohui Zhang , Jing Xin , Hector Garcia de Marina

This paper presents a one-shot learning approach with performance and robustness guarantees for the linear quadratic regulator (LQR) control of stochastic linear systems. Even though data-based LQR control has been widely considered,…

Systems and Control · Electrical Eng. & Systems 2024-10-29 Ramin Esmzad , Hamidreza Modares

We study the problem of \textit{safe control of linear dynamical systems corrupted with non-stochastic noise}, and provide an algorithm that guarantees (i) zero constraint violation of convex time-varying constraints, and (ii) bounded…

Systems and Control · Electrical Eng. & Systems 2023-08-25 Hongyu Zhou , Vasileios Tzoumas

Optimal control of stochastic nonlinear dynamical systems is a major challenge in the domain of robot learning. Given the intractability of the global control problem, state-of-the-art algorithms focus on approximate sequential optimization…

Machine Learning · Computer Science 2020-04-23 Joe Watson , Hany Abdulsamad , Jan Peters
‹ Prev 1 8 9 10 Next ›