English
Related papers

Related papers: Controlling Unknown Linear Dynamics with Almost Op…

200 papers

We consider the problem of online prediction in a marginally stable linear dynamical system subject to bounded adversarial or (non-isotropic) stochastic perturbations. This poses two challenges. Firstly, the system is in general…

Machine Learning · Computer Science 2020-11-24 Udaya Ghai , Holden Lee , Karan Singh , Cyril Zhang , Yi Zhang

We introduce a novel extension of the canonical multi-armed bandit problem that incorporates an additional strategic innovation: abstention. In this enhanced framework, the agent is not only tasked with selecting an arm at each time step,…

Machine Learning · Computer Science 2026-03-24 Junwen Yang , Tianyuan Jin , Vincent Y. F. Tan

We study the problem of online learning with dynamics, where a learner interacts with a stateful environment over multiple rounds. In each round of the interaction, the learner selects a policy to deploy and incurs a cost that depends on…

Machine Learning · Computer Science 2020-12-04 Kush Bhatia , Karthik Sridharan

In online convex optimization, the player aims to minimize regret, or the difference between her loss and that of the best fixed decision in hindsight over the entire repeated game. Algorithms that minimize (standard) regret may converge to…

Machine Learning · Computer Science 2023-02-14 Zhou Lu , Elad Hazan

In the present paper we deal with an optimal control problem related to a model in population dynamics; more precisely, the goal is to modify the behavior of a given density of individuals via another population of agents interacting with…

Optimization and Control · Mathematics 2016-09-26 Mattia Bongini , Giuseppe Buttazzo

Trajectory following is one of the complicated control problems when its dynamics are nonlinear, stochastic and include a large number of parameters. The problem has significant difficulties including a large number of trials required for…

Robotics · Computer Science 2019-02-14 Ali Lenjani

In the convex optimization approach to online regret minimization, many methods have been developed to guarantee a $O(\sqrt{T})$ bound on regret for subdifferentiable convex loss functions with bounded subgradients, by using a reduction to…

Machine Learning · Computer Science 2016-09-20 Arthur Flajolet , Patrick Jaillet

This paper presents a data-driven optimal control policy for a micro flapping wing unmanned aerial vehicle. First, a set of optimal trajectories are computed off-line based on a geometric formulation of dynamics that captures the nonlinear…

Robotics · Computer Science 2022-06-09 Tejaswi K. C. , Taeyoung Lee

Most known regret bounds for reinforcement learning are either episodic or assume an environment without traps. We derive a regret bound without making either assumption, by allowing the algorithm to occasionally delegate an action to an…

Machine Learning · Computer Science 2019-07-22 Vanessa Kosoy

The paper deals with a problem of control of a system characterized by the fact that the influence of controls on the dynamics of certain functions of state variables (called observables) is relatively weak and the rates of change of these…

Optimization and Control · Mathematics 2016-08-12 Vladimir Gaitsgory , Sergey Rossomakhine

This paper is concerned with the design of optimal control for finite-dimensional control-affine nonlinear dynamical systems. We introduce an optimal control problem that specifically optimizes nonlinear observability in addition to…

Systems and Control · Computer Science 2017-08-03 Atiye Alaeddini , Kristi A. Morgansen , Mehran Mesbahi

We present the first computationally-efficient algorithm with $\widetilde O(\sqrt{T})$ regret for learning in Linear Quadratic Control systems with unknown dynamics. By that, we resolve an open question of Abbasi-Yadkori and Szepesv\'ari…

Machine Learning · Computer Science 2019-02-26 Alon Cohen , Tomer Koren , Yishay Mansour

In this paper we consider an optimal control problem governed by a time-dependent variational inequality arising in quasistatic plasticity with linear kinematic hardening. We address certain continuity properties of the forward operator,…

Optimization and Control · Mathematics 2012-09-06 Gerd Wachsmuth

In this paper, we present the combined learning-and-control (CLC) approach, which is a new way to solve optimal control problems with unknown dynamics by unifying model-based control and data-driven learning. The key idea is simple: we…

Systems and Control · Electrical Eng. & Systems 2025-10-02 Panagiotis Kounatidis , Andreas A. Malikopoulos

We consider the problem of discounted optimal state-feedback regulation for general unknown deterministic discrete-time systems. It is well known that open-loop instability of systems, non-quadratic cost functions and complex nonlinear…

Systems and Control · Electrical Eng. & Systems 2020-03-31 Alexandros Tanzanakis , John Lygeros

This note is addressed to giving a short introduction to control theory of stochastic systems, governed by stochastic differential equations in both finite and infinite dimensions. We will mainly explain the new phenomenon and difficulties…

Optimization and Control · Mathematics 2016-12-09 Qi Lu , Xu Zhang

We study Model Predictive Control (MPC) and propose a general analysis pipeline to bound its dynamic regret. The pipeline first requires deriving a perturbation bound for a finite-time optimal control problem. Then, the perturbation bound…

Optimization and Control · Mathematics 2022-10-25 Yiheng Lin , Yang Hu , Guannan Qu , Tongxin Li , Adam Wierman

We propose a learning-based robust predictive control algorithm that compensates for significant uncertainty in the dynamics for a class of discrete-time systems that are nominally linear with an additive nonlinear component. Such systems…

Systems and Control · Electrical Eng. & Systems 2021-10-15 Rohan Sinha , James Harrison , Spencer M. Richards , Marco Pavone

This paper presents a data-driven method to find a closed-loop optimal controller, which minimizes a specified infinite-horizon cost function for systems with unknown dynamics. Suppose the closed-loop optimal controller can be parameterized…

Machine Learning · Computer Science 2025-11-20 Wenjian Hao , Paulo C. Heredia , Shaoshuai Mou

We study online learning in unknown Markov games, a problem that arises in episodic multi-agent reinforcement learning where the actions of the opponents are unobservable. We show that in this challenging setting, achieving sublinear regret…

Machine Learning · Computer Science 2021-02-09 Yi Tian , Yuanhao Wang , Tiancheng Yu , Suvrit Sra
‹ Prev 1 8 9 10 Next ›