中文
相关论文

相关论文: Self-optimizing adaptive optics control with Reinf…

200 篇论文

We propose a new approach for power control in wireless networks using self-supervised learning. We partition a multi-layer perceptron that takes as input the channel matrix and outputs the power control decisions into a backbone and a…

信号处理 · 电气工程与系统科学 2021-02-15 Navid Naderializadeh

This work explores the usage of a supplementary controller for improving the transient performance of inverter$\unicode{x2013}$based resources (IBR) in microgrids. The supplementary controller is trained using a reinforcement learning…

系统与控制 · 电气工程与系统科学 2022-07-12 Ashwin Venkataramanan , Ali Mehrizi-Sani

In this paper, we investigate a design approach of reinforcement learning to engineer a gyroscope in an optical lattice for the inertial sensing of rotations. Our methodology is not based on traditional atom interferometry, that is,…

量子物理 · 物理学 2024-12-02 Liang-Ying Chih , Murray Holland

We propose a reinforcement learning approach for real-time exposure control of a mobile camera that is personalizable. Our approach is based on Markov Decision Process (MDP). In the camera viewfinder or live preview mode, given the current…

计算机视觉与模式识别 · 计算机科学 2018-08-07 Huan Yang , Baoyuan Wang , Noranart Vesdapunt , Minyi Guo , Sing Bing Kang

This paper presents a novel model-reference reinforcement learning control method for uncertain autonomous surface vehicles. The proposed control combines a conventional control method with deep reinforcement learning. With the conventional…

系统与控制 · 电气工程与系统科学 2021-06-17 Qingrui Zhang , Wei Pan , Vasso Reppa

High-temperature superconductors are essential for next-generation energy and quantum technologies, yet their performance is often limited by the critical current density ($J_c$), which is strongly influenced by microstructural defects.…

材料科学 · 物理学 2025-10-28 Mouyang Cheng , Qiwei Wan , Bowen Yu , Eunbi Rha , Michael J Landry , Mingda Li

While pre-trained visual representations have significantly advanced imitation learning, they are often task-agnostic as they remain frozen during policy learning. In this work, we explore leveraging pre-trained text-to-image diffusion…

计算机视觉与模式识别 · 计算机科学 2026-04-09 Heeseong Shin , Byeongho Heo , Dongyoon Han , Seungryong Kim , Taekyung Kim

We have developed a Parallel Integrated Control and Training System, leveraging the deep reinforcement learning to dynamically adjust the control strategies in real time for scanning probe microscopy techniques.

材料科学 · 物理学 2025-02-12 Ziwei Wei , Shuming Wei , Qibin Zeng , Wanheng Lu , Huajun Liu , Kaiyang Zeng

Recently, adversarial imitation learning has shown a scalable reward acquisition method for inverse reinforcement learning (IRL) problems. However, estimated reward signals often become uncertain and fail to train a reliable statistical…

机器学习 · 计算机科学 2023-01-06 Dong-Sig Han , Hyunseo Kim , Hyundo Lee , Je-Hwan Ryu , Byoung-Tak Zhang

Inverse optimal control (IOC) is a promising paradigm for learning and mimicking optimal control strategies from capable demonstrators, or gaining a deeper understanding of their intentions, by estimating an unknown objective function from…

系统与控制 · 电气工程与系统科学 2025-08-28 Rahel Rickenbach , Amon Lahr , Melanie N. Zeilinger

Applying reinforcement learning to robotic systems poses a number of challenging problems. A key requirement is the ability to handle continuous state and action spaces while remaining within a limited time and resource budget.…

机器学习 · 计算机科学 2020-06-29 Benjamin van Niekerk , Andreas Damianou , Benjamin Rosman

Visual-inertial systems rely on precise calibrations of both camera intrinsics and inter-sensor extrinsics, which typically require manually performing complex motions in front of a calibration target. In this work we present a novel…

机器人学 · 计算机科学 2021-02-17 Le Chen , Yunke Ao , Florian Tschopp , Andrei Cramariuc , Michel Breyer , Jen Jen Chung , Roland Siegwart , Cesar Cadena

Recently, safe reinforcement learning (RL) with the actor-critic structure for continuous control tasks has received increasing attention. It is still challenging to learn a near-optimal control policy with safety and convergence…

机器学习 · 计算机科学 2024-02-06 Xinglong Zhang , Yaoqian Peng , Biao Luo , Wei Pan , Xin Xu , Haibin Xie

This paper addresses the inverse optimal control problem of finding the state weighting function that leads to a quadratic value function when the cost on the input is fixed to be quadratic. The paper focuses on a class of infinite horizon…

最优化与控制 · 数学 2022-11-21 Luis Rodrigues

Model predictive control can optimally deal with nonlinear systems under consideration of constraints. The control performance depends on the model accuracy and the prediction horizon. Recent advances propose to use reinforcement learning…

机器学习 · 计算机科学 2024-11-01 Dean Brandner , Sergio Lucia

The success of ground-based instruments for high contrast exoplanet imaging depends on the degree to which adaptive optics (AO) systems can mitigate atmospheric turbulence. While modern AO systems typically suffer from millisecond time lags…

天体物理仪器与方法 · 物理学 2019-09-13 Rebecca Jensen-Clem , Charlotte Z. Bond , Sylvain Cetre , Eden McEwen , Peter Wizinowich , Sam Ragland , Dimitri Mawet , James Graham

The performance of future observatories such as the Extremely Large Telescope is mainly limited by atmospheric turbulence and structural vibrations of the optical assembly. To further enhance the mitigation performance of adaptive optics,…

天体物理仪器与方法 · 物理学 2025-09-17 Pascal Jaufmann , Aaron Buck , Marco Zaiser , Jörg-Uwe Pott , Oliver Sawodny

Although acrobatic flight control has been studied extensively, one key limitation of the existing methods is that they are usually restricted to specific maneuver tasks and cannot change flight pattern parameters online. In this work, we…

机器人学 · 计算机科学 2026-05-18 Zikang Yin , Canlun Zheng , Shiliang Guo , Zhikun Wang , Shiyu Zhao

In this paper a novel model-free algorithm is proposed. This algorithm can learn the nearly optimal control law of constrained-input systems from online data without requiring any a priori knowledge of system dynamics. Based on the concept…

系统与控制 · 电气工程与系统科学 2022-05-03 Han Zhao , Lei Guo

We use Reinforcement Meta-Learning to optimize an adaptive integrated guidance, navigation, and control system suitable for exoatmospheric interception of a maneuvering target. The system maps observations consisting of strapdown seeker…

系统与控制 · 电气工程与系统科学 2021-12-14 Brian Gaudet , Roberto Furfaro , Richard Linares , Andrea Scorsoglio