中文
相关论文

相关论文: Fish Growth Trajectory Tracking via Reinforcement …

200 篇论文

Many applications -- including power systems, robotics, and economics -- involve a dynamical system interacting with a stochastic and hard-to-model environment. We adopt a reinforcement learning approach to control such systems.…

最优化与控制 · 数学 2025-08-26 Abed AlRahman Al Makdah , Oliver Kosut , Lalitha Sankar , Shaofeng Zou

This article develops a methodology that enables learning an objective function of an optimal control system from incomplete trajectory observations. The objective function is assumed to be a weighted sum of features (or basis functions)…

机器人学 · 计算机科学 2021-05-07 Wanxin Jin , Dana Kulić , Shaoshuai Mou , Sandra Hirche

Model-based controllers on real robots require accurate knowledge of the system dynamics to perform optimally. For complex dynamics, first-principles modeling is not sufficiently precise, and data-driven approaches can be leveraged to learn…

机器人学 · 计算机科学 2021-05-17 Weixuan Zhang , Marco Tognon , Lionel Ott , Roland Siegwart , Juan Nieto

We develop a physics-informed learning framework for energy-shaping control of port-Hamiltonian (pH) systems from trajectory data. The proposed approach co-learns a pH system model and an optimal energy-balancing passivity-based controller…

系统与控制 · 电气工程与系统科学 2026-05-07 Ankur Kamboj , Biswadip Dey , Vaibhav Srivastava

This paper demonstrates a data-driven control approach for demand response in real-life residential buildings. The objective is to optimally schedule the heating cycles of the Domestic Hot Water (DHW) buffer to maximize the self-consumption…

系统与控制 · 计算机科学 2017-03-17 Oscar De Somer , Ana Soares , Tristan Kuijpers , Koen Vossen , Koen Vanthournout , Fred Spiessens

Using a model heat engine, we show that neural network-based reinforcement learning can identify thermodynamic trajectories of maximal efficiency. We consider both gradient and gradient-free reinforcement learning. We use an evolutionary…

神经与进化计算 · 计算机科学 2021-12-21 Chris Beeler , Uladzimir Yahorau , Rory Coles , Kyle Mills , Stephen Whitelam , Isaac Tamblyn

The optimal control of a water reservoir systems represents a challenging problem, due to uncertain hydrologic inputs and the need to adapt to changing environment and varying control objectives. In this work, we propose a real-time…

系统与控制 · 电气工程与系统科学 2021-03-26 Pauline Kergus , Simone Formentin , Matteo Giuliani , Andrea Castelletti

Accurately predicting the dynamics of robotic systems is crucial for model-based control and reinforcement learning. The most common way to estimate dynamics is by fitting a one-step ahead prediction model and using it to recursively…

机器学习 · 计算机科学 2021-09-02 Nathan O. Lambert , Albert Wilcox , Howard Zhang , Kristofer S. J. Pister , Roberto Calandra

Learning-based control methods utilize run-time data from the underlying process to improve the controller performance under model mismatch and unmodeled disturbances. This is beneficial for optimizing industrial processes, where the…

系统与控制 · 电气工程与系统科学 2021-11-22 Efe C. Balta , Kira Barton , Dawn M. Tilbury , Alisa Rupenyan , John Lygeros

The foraging behavior of animals is a paradigm of target search in nature. Understanding which foraging strategies are optimal and how animals learn them are central challenges in modeling animal foraging. While the question of optimality…

统计力学 · 物理学 2024-04-15 Gorka Muñoz-Gil , Andrea López-Incera , Lukas J. Fiderer , Hans J. Briegel

Questions remain on the robustness of data-driven learning methods when crossing the gap from simulation to reality. We utilize weight anchoring, a method known from continual learning, to cultivate and fixate desired behavior in Neural…

机器学习 · 计算机科学 2023-04-21 Steffen Gracla , Edgar Beck , Carsten Bockelmann , Armin Dekorsy

We propose a data augmentation method for offline reinforcement learning, motivated by active positioning problems. Particularly, our approach enables the training of off-policy models from a limited number of suboptimal trajectories. We…

机器学习 · 计算机科学 2026-05-14 Tobias Schmähling , Matthias Burkhardt , Tobias Windisch

Imitation learning is a promising approach for enabling generalist capabilities in humanoid robots, but its scaling is fundamentally constrained by the scarcity of high-quality expert demonstrations. This limitation can be mitigated by…

机器人学 · 计算机科学 2025-08-21 Quentin Rouxel , Clemente Donoso , Fei Chen , Serena Ivaldi , Jean-Baptiste Mouret

In this paper, the circle formation control problem is addressed for a group of cooperative underactuated fish-like robots involving unknown nonlinear dynamics and disturbances. Based on the reinforcement learning and cognitive consistency…

机器人学 · 计算机科学 2021-03-10 Tianhao Zhang , Yueheng Li , Shuai Li , Qiwei Ye , Chen Wang , Guangming Xie

This paper studies optimal consensus tracking problem of heterogeneous linear multi-agent systems. By introducing tracking error dynamics, the optimal tracking problem is reformulated as finding a Nash-equilibrium solution of a multi-player…

最优化与控制 · 数学 2019-05-21 Jilie Zhang , Zhanshan Wang , Hongwei Zhang

Guided policy search algorithms have been proven to work with incredible accuracy for not only controlling a complicated dynamical system, but also learning optimal policies from various unseen instances. One assumes true nature of the…

系统与控制 · 电气工程与系统科学 2020-10-02 Prakash Mallick , Zhiyong Chen , Mohsen Zamani

Autonomous modeling of artificial swarms is necessary because manual creation is a time intensive and complicated procedure which makes it impractical. An autonomous approach employing deep reinforcement learning is presented in this study…

机器人学 · 计算机科学 2023-06-09 Suleman Qamar , Saddam Hussain Khan , Muhammad Arif Arshad , Maryam Qamar , Asifullah Khan

We study the fluid dynamics of two fish-like bodies with synchronised swimming patterns. Our studies are based on two-dimensional simulations of viscous incompressible flows. We distinguish between motion patterns that are externally…

Autonomous unpowered flight is a challenge for control and guidance systems: all the energy the aircraft might use during flight has to be harvested directly from the atmosphere. We investigate the design of an algorithm that optimizes the…

机器学习 · 计算机科学 2017-07-19 Erwan Lecarpentier , Sebastian Rapp , Marc Melo , Emmanuel Rachelson

Crop breeding is crucial in improving agricultural productivity while potentially decreasing land usage, greenhouse gas emissions, and water consumption. However, breeding programs are challenging due to long turnover times,…