中文
相关论文

相关论文: Integrating supervised and reinforcement learning …

200 篇论文

Wavefront shaping is a technique for directing light through turbid media. The theoretical aspects of wavefront shaping are well understood, and under near-ideal experimental conditions, accurate predictions for the expected signal…

光学 · 物理学 2024-07-01 Bahareh Mastiani , Daniël W. S. Cox , Ivo M. Vellekoop

Trajectory prediction is critical for autonomous driving, enabling safe and efficient planning in dense, dynamic traffic. Most existing methods optimize prediction accuracy under fixed-length observations. However, real-world driving often…

机器人学 · 计算机科学 2026-03-12 Hao Zhou , Lu Qi , Jason Li , Jie Zhang , Yi Liu , Xu Yang , Mingyu Fan , Fei Luo

One of the greatest challenges in utilizing multimode optical fibers is mode-mixing and inter-modal interference, which scramble the information delivered by the fiber. A common approach for canceling these effects is to tailor the optical…

光学 · 物理学 2020-04-29 Shachar Resisi , Yehonatan Viernik , Sebastien Popoff , Yaron Bromberg

We propose a piecewise learning framework for controlling nonlinear systems with unknown dynamics. While model-based reinforcement learning techniques in terms of some basis functions are well known in the literature, when it comes to more…

最优化与控制 · 数学 2022-04-06 Milad Farsi , Yinan Li , Ye Yuan , Jun Liu

The recent development of novel aerial vehicles capable of physically interacting with the environment leads to new applications such as contact-based inspection. These tasks require the robotic system to exchange forces with…

机器人学 · 计算机科学 2022-07-06 Weixuan Zhang , Lionel Ott , Marco Tognon , Roland Siegwart

In the last few years the concept of an active space telescope has been greatly developed, to meet demanding requirements with a substantial reduction of tolerances, risks and costs. This is the frame of the LATT project (an ESA TRP) and…

We develop a learning-based algorithm for the control of autonomous systems governed by unknown, nonlinear dynamics to satisfy user-specified spatio-temporal tasks expressed as signal temporal logic specifications. Most existing algorithms…

机器人学 · 计算机科学 2021-10-12 Christos K. Verginis , Zhe Xu , Ufuk Topcu

This paper introduces a novel method for self-supervised video representation learning via feature prediction. In contrast to the previous methods that focus on future feature prediction, we argue that a supervisory signal arising from…

计算机视觉与模式识别 · 计算机科学 2020-11-13 Nadine Behrmann , Juergen Gall , Mehdi Noroozi

Applying reinforcement learning to robotic systems poses a number of challenging problems. A key requirement is the ability to handle continuous state and action spaces while remaining within a limited time and resource budget.…

机器学习 · 计算机科学 2020-06-29 Benjamin van Niekerk , Andreas Damianou , Benjamin Rosman

Reinforcement Learning algorithms have recently been proposed to learn time-sequential control policies in the field of autonomous driving. Direct applications of Reinforcement Learning algorithms with discrete action space will yield…

机器学习 · 计算机科学 2019-12-03 Pin Wang , Hanhan Li , Ching-Yao Chan

The wavefront sensors used today at the biggest World's telescopes have either a high dynamic range or a high sensitivity, and they are subject to a linear trade off between these two parameters. A new class of wavefront sensors, the…

Continuous wavefront sensing on future space telescopes allows relaxation of stability requirements while still allowing on-orbit diffraction-limited optical performance. We consider the suitability of phase retrieval to continuously…

天体物理仪器与方法 · 物理学 2023-09-14 Hyukmo Kang , Kyle Van Gorkom , Jess Johnson , Ole Singlestad , Aaron Goldtooth , Daewook Kim , Ewan S. Douglas

In this work, we present an approach to supervisory reinforcement learning control for unmanned aerial vehicles (UAVs). UAVs are dynamic systems where control decisions in response to disturbances in the environment have to be made in the…

系统与控制 · 电气工程与系统科学 2023-05-23 Ibrahim Ahmed , Marcos Quinones-Grueiro , Gautam Biswas

There has recently been an increased interest in reinforcement learning for nonlinear control problems. However standard reinforcement learning algorithms can often struggle even on seemingly simple set-point control problems. This paper…

系统与控制 · 电气工程与系统科学 2023-04-21 Ruoqi Zhang , Per Mattsson , Torbjörn Wigren

Astronomical adaptive optics systems with open-loop deformable mirror control have recently come on-line. In these systems, the deformable mirror surface is not included in the wavefront sensor paths, and so changes made to the deformable…

天体物理仪器与方法 · 物理学 2015-05-30 Alastair Basden , Richard Myers , Eric Gendron

Online model predictive control (MPC) for piecewise affine (PWA) systems requires the online solution to an optimization problem that implicitly optimizes over the switching sequence of PWA regions, for which the computational burden can be…

系统与控制 · 电气工程与系统科学 2025-03-27 Samuel Mallick , Azita Dabiri , Bart De Schutter

This paper puts forth a new, reconfigurable intelligent surface (RIS)-assisted, uplink, user-centric cell-free (UCCF) system managed with the assistance of a digital twin (DT). Specifically, we propose a novel learning framework that…

信号处理 · 电气工程与系统科学 2023-02-13 Yingping Cui , Tiejun Lv , Wei Ni , Abbas Jamalipour

We introduce a Transformer-based Reinforcement Learning framework for autonomous orbital collision avoidance that explicitly models the effects of partial observability and imperfect monitoring in space operations. The framework combines a…

机器学习 · 计算机科学 2026-03-26 Thomas Georges , Adam Abdin

The development of low-loss reconfigurable integrated optical devices enables further research into technologies including photonic signal processing, analogue quantum computing, and optical neural networks. Here, we introduce digital…

Wavefront sensing and control are important for enabling one of the key advantages of using large apertures, namely higher angular resolutions. Pyramid wavefront sensors are becoming commonplace in new instrument designs owing to their…

天体物理仪器与方法 · 物理学 2019-03-20 Julien Lozi , Nemanja Jovanovic , Olivier Guyon , Mark Chun , Shane Jacobson , Sean Goebel , Frantz Martinache