中文
相关论文

相关论文: Self-optimizing adaptive optics control with Reinf…

200 篇论文

The optimal predictor for a linear dynamical system (with hidden state and Gaussian noise) takes the form of an autoregressive linear filter, namely the Kalman filter. However, a fundamental problem in reinforcement learning and control…

机器学习 · 计算机科学 2019-05-27 Holden Lee , Cyril Zhang

The behavior of an adaptive optics (AO) system for ground-based high contrast imaging (HCI) dictates the achievable contrast of the instrument. In conditions where the coherence time of the atmosphere is short compared to the speed of the…

天体物理仪器与方法 · 物理学 2021-08-23 Maaike A. M. van Kooten , Rebecca Jensen-Clem , Sylvain Cetre , Sam Ragland , Charlotte Z. Bond , J. Fowler , Peter Wizinowich

The direct imaging of potentially habitable exoplanets is one prime science case for high-contrast imaging instruments on extremely large telescopes. Most such exoplanets orbit close to their host stars, where their observation is limited…

天体物理仪器与方法 · 物理学 2026-05-27 Jalo Nousiainen , Iremsu Taskin , Markus Kasper , Gilles Orban De Xivry , Olivier Absil

Adjusting camera exposure in arbitrary lighting conditions is the first step to ensure the functionality of computer vision applications. Poorly adjusted camera exposure often leads to critical failure and performance degradation.…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Kyunghyun Lee , Ukcheol Shin , Byeong-Uk Lee

This paper proposes a robust control design method using reinforcement-learning for controlling partially-unknown dynamical systems under uncertain conditions. The method extends the optimal reinforcement-learning algorithm with a new…

系统与控制 · 电气工程与系统科学 2020-04-17 Phuong D. Ngo , Fred Godtliebsen

Common approaches to control a data-center cooling system rely on approximated system/environment models that are built upon the knowledge of mechanical cooling and electrical and thermal management. These models are difficult to design and…

系统与控制 · 计算机科学 2018-08-31 Takao Moriyama , Giovanni De Magistris , Michiaki Tatsubori , Tu-Hoa Pham , Asim Munawar , Ryuki Tachibana

This paper proposes a novel robust reinforcement learning framework for discrete-time linear systems with model mismatch that may arise from the sim-to-real gap. A key strategy is to invoke advanced techniques from control theory. Using the…

系统与控制 · 电气工程与系统科学 2023-12-07 Leilei Cui , Tamer Başar , Zhong-Ping Jiang

We present an orientation adaptive controller to compensate for the effects of highly constrained environments on continuum manipulator actuation. A transformation matrix updated using optimal estimation techniques from optical flow…

机器人学 · 计算机科学 2019-09-04 Mrinal Verghese , Florian Richter , Aaron Gunn , Phil Weissbrod , Michael Yip

This paper proposes a reinforcement learning approach for traffic control with the adaptive horizon. To build the controller for the traffic network, a Q-learning-based strategy that controls the green light passing time at the network…

系统与控制 · 计算机科学 2019-04-01 Wentao Chen , Tehuan Chen , Guang Lin

This paper proposes a reinforcement learning-based approach for optimal transient frequency control in power systems with stability and safety guarantees. Building on Lyapunov stability theory and safety-critical control, we derive…

系统与控制 · 电气工程与系统科学 2024-02-22 Zhenyi Yuan , Changhong Zhao , Jorge Cortes

The behavior of an adaptive optics (AO) system for ground-based high contrast imaging (HCI) dictates the achievable contrast of the instrument. In conditions where the coherence time of the atmosphere is short compared to the speed of the…

This paper introduces a novel model-free and a partially model-free algorithm for inverse optimal control (IOC), also known as inverse reinforcement learning (IRL), aimed at estimating the cost function of continuous-time nonlinear…

系统与控制 · 电气工程与系统科学 2025-03-20 Hamed Jabbari Asl , Eiji Uchibe

Contrastive representation learning has been recently proved to be very efficient for self-supervised training. These methods have been successfully used to train encoders which perform comparably to supervised training on downstream…

机器学习 · 计算机科学 2020-12-03 Ibrahim Merad , Yiyang Yu , Emmanuel Bacry , Stéphane Gaïffas

Time-delay error is a significant error source in adaptive optics (AO) systems. It arises from the latency between sensing the wavefront and applying the correction. Predictive control algorithms reduce the time-delay error, providing…

天体物理仪器与方法 · 物理学 2024-06-27 Jalo Nousiainen , Juha-Pekka Puska , Tapio Helin , Nuutti Hyvönen , Markus Kasper

In this paper, we propose a model-free adaptive learning solution for a model-following control problem. This approach employs policy iteration, to find an optimal adaptive control solution. It utilizes a moving finite-horizon of…

系统与控制 · 电气工程与系统科学 2023-02-07 Mohammed I. Abouheaf , Hashim A. Hashim , Mohammad A. Mayyas , Kyriakos G. Vamvoudakis

For future extremely large telescopes, error in extreme adaptive optics systems at small angular separations will be highly impacted by the lag time of the correction, which is typically on millisecond timescales; one solution is to apply a…

天体物理仪器与方法 · 物理学 2023-10-05 J. Fowler , M. A. M. van Kooten , R. Jensen-Clem

Reinforcement learning was carried out in a simulated environment to learn continuous velocity control over multiple motor axes. This was then applied to a real-world optical tweezers experiment with the objective of moving a laser-trapped…

机器学习 · 计算机科学 2020-11-11 Matthew Praeger , Yunhui Xie , James A. Grant-Jacob , Robert W. Eason , Ben Mills

We present theoretical and numerical results concerning the problem to find the path that minimizes the time to navigate between two given points in a complex fluid under realistic navigation constraints. We contrast deterministic Optimal…

系统与控制 · 电气工程与系统科学 2021-03-02 Michele Buzzicotti , Luca Biferale , Fabio Bonaccorso , Patricio Clark di Leoni , Kristian Gustavsson

Robotic systems that rely primarily on self-supervised learning have the potential to decrease the amount of human annotation and engineering effort required to learn control strategies. In the same way that prior robotic systems have…

This paper investigates the use of Reinforcement Learning for the robust design of low-thrust interplanetary trajectories in presence of severe disturbances, modeled alternatively as Gaussian additive process noise, observation noise,…

机器学习 · 计算机科学 2020-08-20 Alessandro Zavoli , Lorenzo Federici