中文
相关论文

相关论文: Development of a PPO-Reinforcement Learned Walking…

200 篇论文

Proximal Policy Optimization (PPO) has been positioned by recent literature as the canonical method for the RL part of Reinforcement Learning from Human Feedback (RLHF). PPO performs well empirically but has a heuristic motivation and…

机器学习 · 计算机科学 2026-02-10 Dipan Maity

Proximal Policy Optimization (PPO) is a popular model-free reinforcement learning algorithm, esteemed for its simplicity and efficacy. However, due to its inherent on-policy nature, its proficiency in harnessing data from disparate policies…

机器学习 · 计算机科学 2024-06-07 Yaozhong Gan , Renye Yan , Xiaoyang Tan , Zhe Wu , Junliang Xing

Legged robots have the potential to become vital in maintenance, home support, and exploration scenarios. In order to interact with and manipulate their environments, most legged robots are equipped with a dedicated robot arm, which means…

机器人学 · 计算机科学 2024-02-19 Philip Arm , Mayank Mittal , Hendrik Kolvenbach , Marco Hutter

Locomotion has seen dramatic progress for walking or running across challenging terrains. However, robotic quadrupeds are still far behind their biological counterparts, such as dogs, which display a variety of agile skills and can use the…

机器人学 · 计算机科学 2023-03-23 Xuxin Cheng , Ashish Kumar , Deepak Pathak

In this paper, with a view toward deployment of light-weight control frameworks for bipedal walking robots, we realize end-foot trajectories that are shaped by a single linear feedback policy. We learn this policy via a model-free and a…

机器人学 · 计算机科学 2021-08-10 Lokesh Krishna , Utkarsh A. Mishra , Guillermo A. Castillo , Ayonga Hereid , Shishir Kolathaya

Legged robots, specifically quadrupeds, are becoming increasingly attractive for industrial applications such as inspection. However, to leave the laboratory and to become useful to an end user requires reliability in harsh conditions. From…

机器人学 · 计算机科学 2019-08-13 David Wisth , Marco Camurri , Maurice Fallon

With the rapid development of embodied intelligence, locomotion control of quadruped robots on complex terrains has become a research hotspot. Unlike traditional locomotion control approaches focusing solely on velocity tracking, we pursue…

机器人学 · 计算机科学 2025-03-07 Xiangyu Miao , Jun Sun , Hang Lai , Xinpeng Di , Jiahang Cao , Yong Yu , Weinan Zhang

Legged locomotion on flowing ground ({\em e.g.} granular media) is unlike locomotion on hard ground because feet experience both solid- and fluid-like forces during surface penetration. Recent bio-inspired legged robots display speed…

A wide range of microorganisms, e.g. bacteria, propel themselves by rotation of soft helical tails, also known as flagella. Due to the small size of these organisms, viscous forces overwhelm inertial effects and the flow is at low Reynolds…

机器人学 · 计算机科学 2021-03-11 Yayun Du , Andrew Miller , Mohammad Khalid Jawed

This paper describes the hardware, software framework, and experimental testing of SURENA IV humanoid robotics platform. SURENA IV has 43 degrees of freedom (DoFs), including seven DoFs for each arm, six DoFs for each hand, and six DoFs for…

SLOT (Soft Legged Omnidirectional Tetrapod), a tendon-driven soft quadruped robot with 3D-printed TPU legs, is presented to study physics-informed modeling and control of compliant legged locomotion using only four actuators. Each leg is…

机器人学 · 计算机科学 2026-02-19 Saumya Karan , Neerav Maram , Suraj Borate , Madhu Vadali

Reinforcement learning (RL) is a promising approach for solving robotic manipulation tasks. However, it is challenging to apply the RL algorithms directly in the real world. For one thing, RL is data-intensive and typically requires…

机器人学 · 计算机科学 2026-04-24 Weirui Ye , Yunsheng Zhang , Haoyang Weng , Xianfan Gu , Shengjie Wang , Tong Zhang , Mengchen Wang , Pieter Abbeel , Yang Gao

This paper presents a control framework to teleoperate a quadruped robot's foot for operator-guided haptic exploration of the environment. Since one leg of a quadruped robot typically only has 3 actuated degrees of freedom (DoFs), the torso…

机器人学 · 计算机科学 2020-10-27 Guiyang Xin , Joshua Smith , David Rytz , Wouter Wolfslag , Hsiu-Chin Lin , Michael Mistry

This study investigates formal-method-based trajectory optimization (TO) for bipedal locomotion, focusing on scenarios where the robot encounters external perturbations at unforeseen times. Our key research question centers around the…

机器人学 · 计算机科学 2023-10-18 Zhaoyuan Gu , Rongming Guo , William Yates , Yipu Chen , Ye Zhao

Developing robot controllers capable of achieving dexterous nonprehensile manipulation, such as pushing an object on a table, is challenging. The underactuated and hybrid-dynamics nature of the problem, further complicated by the…

机器人学 · 计算机科学 2023-08-07 Juan Del Aguila Ferrandis , João Moura , Sethu Vijayakumar

Whole-body manipulation is a powerful yet underexplored approach that enables robots to interact with large, heavy, or awkward objects using more than just their end-effectors. Soft robots, with their inherent passive compliance, are…

机器人学 · 计算机科学 2025-09-30 Curtis C. Johnson , Carlo Alessi , Egidio Falotico , Marc D. Killpack

Flexible manufacturing processes demand robots to easily adapt to changes in the environment and interact with humans. In such dynamic scenarios, robotic tasks may be programmed through learning-from-demonstration approaches, where a…

机器人学 · 计算机科学 2019-08-21 Leonel Rozo

Vision-based robotic cloth unfolding has made great progress recently. However, prior works predominantly rely on value learning and have not fully explored policy-based techniques. Recently, the success of reinforcement learning on the…

计算机视觉与模式识别 · 计算机科学 2024-05-09 Libing Yang , Yang Li , Long Chen

Robot person following (RPF) -- mobile robots that follow and assist a specific person -- has emerging applications in personal assistance, security patrols, eldercare, and logistics. To be effective, such robots must follow the target…

机器人学 · 计算机科学 2026-05-14 Hanjing Ye , Weixi Situ , Jianwei Peng , Yu Zhan , Bingyi Xia , Kuanqi Cai , Hong Zhang

Safe exploration is a key to applying reinforcement learning (RL) in safety-critical systems. Existing safe exploration methods guaranteed safety under the assumption of regularity, and it has been difficult to apply them to large-scale…

机器学习 · 计算机科学 2021-11-10 Akifumi Wachi , Yunyue Wei , Yanan Sui
‹ 上一页 1 8 9 10 下一页 ›