中文
相关论文

相关论文: Continuous Control with Contexts, Provably

200 篇论文

An important feature of a dynamic game is its monitoring structure namely, what the players effectively see from the played actions. We consider games with arbitrary monitoring structures. One of the purposes of this paper is to know to…

信息论 · 计算机科学 2012-10-24 Maël Le Treust , Samson Lasaulce

This paper proposes a simulation-based reinforcement learning algorithm for controlling systems with uncertain and varying system parameters. While simulators are useful for safely learning control policies, the reality gap remains a major…

系统与控制 · 电气工程与系统科学 2026-05-14 Junya Ikemoto

Deep reinforcement learning has the potential to address various scientific problems. In this paper, we implement an optics simulation environment for reinforcement learning based controllers. The environment captures the essence of…

机器学习 · 计算机科学 2023-10-03 Abulikemu Abuduweili , Changliu Liu

Many physical AI tasks are governed by implicit equilibrium: an agent actuates a subset of degrees of freedom (boundary DoFs), while the remaining free DoFs settle by minimizing a total potential energy. Even seemingly basic tasks such as…

机器人学 · 计算机科学 2026-05-06 Dezhong Tong , Jiawen Wang , Hengyi Zhou , Yinglong Shen , Xiaonan Huang , M. Khalid Jawed

We consider what we call the offline-to-online learning setting, focusing on stochastic finite-armed bandit problems. In offline-to-online learning, a learner starts with offline data collected from interactions with an unknown environment…

机器学习 · 计算机科学 2025-03-11 Flore Sentenac , Ilbin Lee , Csaba Szepesvari

In modern machine learning, models can often fit training data in numerous ways, some of which perform well on unseen (test) data, while others do not. Remarkably, in such cases gradient descent frequently exhibits an implicit bias that…

机器学习 · 计算机科学 2024-06-04 Noam Razin , Yotam Alexander , Edo Cohen-Karlik , Raja Giryes , Amir Globerson , Nadav Cohen

Data-driven control in unknown environments requires a clear understanding of the involved uncertainties for ensuring safety and efficient exploration. While aleatoric uncertainty that arises from measurement noise can often be explicitly…

机器学习 · 计算机科学 2023-07-13 Neha Das , Jonas Umlauft , Armin Lederer , Thomas Beckers , Sandra Hirche

Climate Change is an incredibly complicated problem that humanity faces. When many variables interact with each other, it can be difficult for humans to grasp the causes and effects of the very large-scale problem of climate change. The…

机器学习 · 计算机科学 2022-12-01 Theodore Wolf

We study the linear quadratic Gaussian (LQG) control problem, in which the controller's observation of the system state is such that a desired cost is unattainable. To achieve the desired LQG cost, we introduce a communication link from the…

最优化与控制 · 数学 2021-09-28 Oron Sabag , Peida Tian , Victoria Kostina , Babak Hassibi

We provide an algorithm for the simultaneous system identification and model predictive control of nonlinear systems. The algorithm has finite-time near-optimality guarantees and asymptotically converges to the optimal (non-causal)…

机器人学 · 计算机科学 2025-11-04 Hongyu Zhou , Vasileios Tzoumas

Distributed optimal control is known to be challenging and can become intractable even for linear-quadratic regulator problems. In this work, we study a special class of such problems where distributed state feedback controllers can give…

系统与控制 · 电气工程与系统科学 2024-03-14 Johan Olsson , Runyu Zhang , Emma Tegling , Na Li

Standard model-based control design deteriorates when the system dynamics change during operation. To overcome this challenge, online and adaptive methods have been proposed in the literature. In this work, we consider the class of…

系统与控制 · 电气工程与系统科学 2026-04-16 Marcell Bartos , Johannes Köhler , Florian Dörfler , Melanie N. Zeilinger

Consider a linear quadratic regulator (LQR) problem being solved in a model-free manner using the policy gradient approach. If the gradient of the quadratic cost is being transmitted across a rate-limited channel, both the convergence and…

最优化与控制 · 数学 2024-09-20 Lintao Ye , Aritra Mitra , Vijay Gupta

We consider the classic stochastic linear quadratic regulator (LQR) problem under an infinite horizon average stage cost. By leveraging recent policy gradient methods from reinforcement learning, we obtain a first-order method that finds a…

最优化与控制 · 数学 2025-02-21 Caleb Ju , Georgios Kotsalis , Guanghui Lan

Artificial agents can achieve strong task performance while remaining opaque with respect to internal regulation, uncertainty management, and stability under stochastic perturbation. We present IRAM-Omega-Q, a computational architecture…

人工智能 · 计算机科学 2026-03-18 Veronique Ziegler

Motion synthesis in a dynamic environment has been a long-standing problem for character animation. Methods using motion capture data tend to scale poorly in complex environments because of their larger capturing and labeling requirement.…

机器学习 · 计算机科学 2021-01-06 Ying-Sheng Luo , Jonathan Hans Soeseno , Trista Pei-Chun Chen , Wei-Chao Chen

This paper considers optimal control of a quadrotor unmanned aerial vehicles (UAV) using the discrete-time, finite-horizon, linear quadratic regulator (LQR). The state of a quadrotor UAV is represented as an element of the matrix Lie group…

机器人学 · 计算机科学 2021-05-31 Mitchell R. Cohen , Khairi Abdulrahim , James Richard Forbes

We consider policy gradient algorithms for the indefinite least squares stationary optimal control, e.g., linear-quadratic-regulator (LQR) with indefinite state and input penalization matrices. Such a setup has important applications in…

最优化与控制 · 数学 2020-02-13 Jingjing Bu , Mehran Mesbahi

The quadrotor unmanned aerial vehicle is a great platform for control systems research as its nonlinear nature and under-actuated configuration make it ideal to synthesize and analyze control algorithms. After a brief explanation of the…

系统与控制 · 计算机科学 2016-02-09 Andrew Zulu , Samuel John

We consider the problem of online adaptive control of the linear quadratic regulator, where the true system parameters are unknown. We prove new upper and lower bounds demonstrating that the optimal regret scales as…

机器学习 · 计算机科学 2023-10-05 Max Simchowitz , Dylan J. Foster
‹ 上一页 1 8 9 10 下一页 ›