中文
相关论文

相关论文: Optimising low-Reynolds-number predation via optim…

200 篇论文

Offline reinforcement learning (RL) aims to learn a policy that maximizes the expected return using a given static dataset of transitions. However, offline RL faces the distribution shift problem. The policy constraint offline RL method is…

机器学习 · 计算机科学 2025-12-24 Yuanhao Chen , Qi Liu , Pengbin Chen , Zhongjian Qiao , Yanjie Li

Compact quadrupedal robots are proving increasingly suitable for deployment in real-world scenarios. Their smaller size fosters easy integration into human environments. Nevertheless, real-time locomotion on uneven terrains remains…

机器人学 · 计算机科学 2026-02-20 Davide Plozza , Patricia Apostol , Paul Joseph , Simon Schläpfer , Michele Magno

We address the swimming problem at low Reynolds number. This regime, which is typically used for micro-swimmers, is described by Stokes equations. We couple a PDE solver of Stokes equations, derived from the Feel++ finite elements library,…

偏微分方程分析 · 数学 2019-11-28 Luca Berti , Laetitia Giraldi , Christophe Prud'Homme

A random recurrent neural network, called a reservoir, can be used to learn robot movements conditioned on context inputs that encode task goals. The Learning is achieved by mapping the random dynamics of the reservoir modulated by context…

机器人学 · 计算机科学 2024-11-19 Zahra Koulaeizadeh , Erhan Oztop

Reinforcement learning methods as a promising technique have achieved superior results in the motion planning of free-floating space robots. However, due to the increase in planning dimension and the intensification of system dynamics…

机器人学 · 计算机科学 2022-09-07 Yuxue Cao , Shengjie Wang , Xiang Zheng , Wenke Ma , Xinru Xie , Lei Liu

Despite some successful applications of goal-driven navigation, existing deep reinforcement learning (DRL)-based approaches notoriously suffers from poor data efficiency issue. One of the reasons is that the goal information is decoupled…

机器人学 · 计算机科学 2023-11-09 Wenhui Huang , Yanxin Zhou , Xiangkun He , Chen Lv

Jumping constitutes an essential component of quadruped robots' locomotion capabilities, which includes dynamic take-off and adaptive landing. Existing quadrupedal jumping studies mainly focused on the stance and flight phase by assuming a…

机器人学 · 计算机科学 2025-09-17 Renjie Wang , Shangke Lyu , Xin Lang , Wei Xiao , Donglin Wang

Recent advances in reinforcement learning have demonstrated its ability to solve hard agent-environment interaction tasks on a super-human level. However, the application of reinforcement learning methods to practical and real-world tasks…

人工智能 · 计算机科学 2021-12-03 Oleg Svidchenko , Aleksei Shpilman

Understanding the complex patterns in space-time exhibited by active systems has been the subject of much interest in recent times. Complementing this forward problem is the inverse problem of controlling active matter. Here we use optimal…

软凝聚态物质 · 物理学 2022-10-12 Suraj Shankar , Vidya Raju , L. Mahadevan

An investigation of optimal feedback controllers' performance and robustness is carried out for vortex shedding behind a 2D cylinder at low Reynolds numbers. To facilitate controller design, we present an efficient modelling approach in…

流体动力学 · 物理学 2020-07-15 Bo Jin , Simon J. Illingworth , Richard D. Sandberg

Many high-performance human activities are executed with little or no external feedback: think of a figure skater landing a triple jump, a pitcher throwing a curveball for a strike, or a barista pouring latte art. To study the process of…

人工智能 · 计算机科学 2025-12-10 Antonio Terpin , Raffaello D'Andrea

Offline reinforcement learning aims to learn from pre-collected datasets without active exploration. This problem faces significant challenges, including limited data availability and distributional shifts. Existing approaches adopt a…

机器学习 · 计算机科学 2024-10-01 Yue Wang , Jinjun Xiong , Shaofeng Zou

Continuous normalizing flows (CNFs) construct invertible mappings between an arbitrary complex distribution and an isotropic Gaussian distribution using Neural Ordinary Differential Equations (neural ODEs). It has not been tractable on…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Shian Du , Yihong Luo , Wei Chen , Jian Xu , Delu Zeng

Trajectory optimization (TO) is an efficient tool to generate a redundant manipulator's joint trajectory following a 6-dimensional Cartesian path. The optimization performance largely depends on the quality of initial trajectories. However,…

机器人学 · 计算机科学 2026-02-10 Minsung Yoon , Mincheul Kang , Daehyung Park , Sung-Eui Yoon

Meta-Bayesian optimisation (meta-BO) aims to improve the sample efficiency of Bayesian optimisation by leveraging data from related tasks. While previous methods successfully meta-learn either a surrogate model or an acquisition function…

机器学习 · 计算机科学 2023-12-25 Alexandre Maraval , Matthieu Zimmer , Antoine Grosnit , Haitham Bou Ammar

This paper presents a novel algorithm for the continuous control of dynamical systems that combines Trajectory Optimization (TO) and Reinforcement Learning (RL) in a single framework. The motivations behind this algorithm are the two main…

For robots to handle the numerous factors that can affect them in the real world, they must adapt to changes and unexpected events. Evolutionary robotics tries to solve some of these issues by automatically optimizing a robot for a specific…

机器人学 · 计算机科学 2018-05-10 Tønnes F. Nygaard , Charles P. Martin , Eivind Samuelsen , Jim Torresen , Kyrre Glette

We propose ReinFlow, a simple yet effective online reinforcement learning (RL) framework that fine-tunes a family of flow matching policies for continuous robotic control. Derived from rigorous RL theory, ReinFlow injects learnable noise…

机器人学 · 计算机科学 2026-01-09 Tonghe Zhang , Chao Yu , Sichang Su , Yu Wang

We address online combinatorial optimization when the player has a prior over the adversary's sequence of losses. In this framework, Russo and Van Roy proposed an information-theoretic analysis of Thompson Sampling based on the information…

机器学习 · 计算机科学 2022-04-05 Sébastien Bubeck , Mark Sellke

We present a novel reinforcement learning (RL) environment designed to both optimize industrial sorting systems and study agent behavior in evolving spaces. In simulating material flow within a sorting process our environment follows the…

机器学习 · 计算机科学 2025-03-14 Tom Maus , Nico Zengeler , Tobias Glasmachers