中文
相关论文

相关论文: Model-Free versus Model-Based Reinforcement Learni…

200 篇论文

Real-time adaptation is imperative to the control of robots operating in complex, dynamic environments. Adaptive control laws can endow even nonlinear systems with good trajectory tracking performance, provided that any uncertain dynamics…

机器人学 · 计算机科学 2022-04-15 Spencer M. Richards , Navid Azizan , Jean-Jacques Slotine , Marco Pavone

Model-free reinforcement learning algorithms can compute policy gradients given sampled environment transitions, but require large amounts of data. In contrast, model-based methods can use the learned model to generate new data, but model…

机器学习 · 计算机科学 2022-03-04 Lukas P. Fröhlich , Maksym Lefarov , Melanie N. Zeilinger , Felix Berkenkamp

Event-triggered communication and control provide high control performance in networked control systems without overloading the communication network. However, most approaches require precise mathematical models of the system dynamics,…

系统与控制 · 电气工程与系统科学 2023-05-16 Lukas Kesper , Sebastian Trimpe , Dominik Baumann

Terrestrial and aerial bimodal vehicles have gained widespread attention due to their cross-domain maneuverability. Nevertheless, their bimodal dynamics significantly increase the complexity of motion planning and control, thus hindering…

机器人学 · 计算机科学 2024-03-04 Ruibin Zhang , Junxiao Lin , Yuze Wu , Yuman Gao , Chi Wang , Chao Xu , Yanjun Cao , Fei Gao

Reinforcement learning has been established over the past decade as an effective tool to find optimal control policies for dynamical systems, with recent focus on approaches that guarantee safety during the learning and/or execution phases.…

系统与控制 · 电气工程与系统科学 2021-10-06 S M Nahid Mahmud , Scott A Nivison , Zachary I. Bell , Rushikesh Kamalapurkar

Reinforcement learning algorithms have shown great success in solving different problems ranging from playing video games to robotics. However, they struggle to solve delicate robotic problems, especially those involving contact…

机器人学 · 计算机科学 2020-07-15 Miroslav Bogdanovic , Majid Khadiv , Ludovic Righetti

Airborne Wind Energy is a lightweight technology that allows power extraction from the wind using airborne devices such as kites and gliders, where the airfoil orientation can be dynamically controlled in order to maximize performance. The…

流体动力学 · 物理学 2022-03-29 N. Orzan , C. Leone , A. Mazzolini , J. Oyero , A. Celani

Designing missiles' autopilot controllers has been a complex task, given the extensive flight envelope and the nonlinear flight dynamics. A solution that can excel both in nominal performance and in robustness to uncertainties is still to…

机器学习 · 计算机科学 2021-09-21 Bernardo Cortez

Proportional-Integrator-Derivative (PID) controller is used in a wide range of industrial and experimental processes. There are a couple of offline methods for tuning PID gains. However, due to the uncertainty of model parameters and…

系统与控制 · 电气工程与系统科学 2025-08-19 Iman Sharifi , Aria Alasty

This study presents an Actor-Critic reinforcement learning Compensated Model Predictive Controller (AC2MPC) designed for high-speed, off-road autonomous driving on deformable terrains. Addressing the difficulty of modeling unknown…

机器人学 · 计算机科学 2026-01-23 Prakhar Gupta , Jonathon M. Smereka , Yunyi Jia

This paper addresses the boundary stabilization of a flexible wing model, both in bending and twisting displacements, under unsteady aerodynamic loads, and in presence of a store. The wing dynamics is captured by a distributed parameter…

最优化与控制 · 数学 2018-08-29 Hugo Lhachemi , David Saussié , Guchuan Zhu

A substantial part of fighter pilot training is simulation-based and involves computer-generated forces controlled by predefined behavior models. The behavior models are typically manually created by eliciting knowledge from experienced…

机器学习 · 计算机科学 2025-10-20 Andreas Strand , Patrick Gorton , Martin Asprusten , Karsten Brathen

It is doubtful that animals have perfect inverse models of their limbs (e.g., what muscle contraction must be applied to every joint to reach a particular location in space). However, in robot control, moving an arm's end-effector to a…

机器人学 · 计算机科学 2022-09-19 Justus Huebotter , Serge Thill , Marcel van Gerven , Pablo Lanillos

Linear dynamical systems that obey stochastic differential equations are canonical models. While optimal control of known systems has a rich literature, the problem is technically hard under model uncertainty and there are hardly any…

系统与控制 · 电气工程与系统科学 2023-06-09 Mohamad Kazem Shirani Faradonbeh , Mohamad Sadegh Shirani Faradonbeh

Although evidence integration to the boundary model has successfully explained a wide range of behavioral and neural data in decision making under uncertainty, how animals learn and optimize the boundary remains unresolved. Here, we propose…

神经与进化计算 · 计算机科学 2024-08-13 Jamal Esmaily , Rani Moran , Yasser Roudi , Bahador Bahrami

Autonomous off-road driving is challenging as risky actions taken by the robot may lead to catastrophic damage. As such, developing controllers in simulation is often desirable as it provides a safer and more economical alternative.…

机器人学 · 计算机科学 2023-10-16 Sean J. Wang , Honghao Zhu , Aaron M. Johnson

In recent times, reinforcement learning has produced baffling results when it comes to performing control tasks with highly non-linear systems. The impressive results always outweigh the potential vulnerabilities or uncertainties associated…

机器人学 · 计算机科学 2023-11-14 Arshad Javeed

Despite substantial growth in wind energy technology in recent decades, aerodynamic modeling of wind turbines relies on momentum models derived in the late 19th and early 20th centuries, which are well-known to break down under flow regimes…

流体动力学 · 物理学 2024-01-19 Jaime Liew , Kirby S. Heck , Michael F. Howland

Carrier landing of aircrafts is a challenge for control due to the existence of nonlinear wind disturbances and the requirements of changing reference trajectories. In this paper, a robust landing control system is presented for carrier…

系统与控制 · 电气工程与系统科学 2025-02-03 Mikhail Kistyarev , Xinhua Wang

We apply a reinforcement meta-learning framework to optimize an integrated and adaptive guidance and flight control system for an air-to-air missile. The system is implemented as a policy that maps navigation system outputs directly to…

系统与控制 · 电气工程与系统科学 2022-05-05 Brian Gaudet , Roberto Furfaro