中文
相关论文

相关论文: End-to-end Deep Reinforcement Learning for Stochas…

200 篇论文

In this paper, a novel deep reinforcement learning (DRL)-based method is proposed to navigate the robot team through unknown complex environments, where the geometric centroid of the robot team aims to reach the goal position while avoiding…

机器人学 · 计算机科学 2019-07-04 Juntong Lin , Xuyun Yang , Peiwei Zheng , Hui Cheng

Unmanned Surface Vehicles technology (USVs) is an exciting topic that essentially deploys an algorithm to safely and efficiently performs a mission. Although reinforcement learning is a well-known approach to modeling such a task,…

机器学习 · 计算机科学 2020-03-24 Mohammad Etemad , Nader Zare , Mahtab Sarvmaili , Amilcar Soares , Bruno Brandoli Machado , Stan Matwin

In the last years, there has been a great interest in machine-learning-based heuristics for solving NP-hard combinatorial optimization problems. The developed methods have shown potential on many optimization problems. In this paper, we…

最优化与控制 · 数学 2022-12-19 Mouad Morabit , Guy Desaulniers , Andrea Lodi

This article presents a deep reinforcement learning-based approach to tackle a persistent surveillance mission requiring a single unmanned aerial vehicle initially stationed at a depot with fuel or time-of-flight constraints to repeatedly…

机器人学 · 计算机科学 2024-05-06 Manav Mishra , Hritik Bana , Saswata Sarkar , Sujeevraja Sanjeevi , PB Sujit , Kaarthik Sundar

A novel hierarchical Deep Neural Network (DNN) model is presented to address the task of end-to-end driving. The model consists of a master classifier network which determines the driving task required from an input stereo image and directs…

机器学习 · 计算机科学 2020-12-03 Jose Solomon , Francois Charette

Reinforcement learning is considered to be a strong AI paradigm which can be used to teach machines through interaction with the environment and learning from their mistakes, but it has not yet been successfully used for automotive…

机器学习 · 统计学 2016-12-14 Ahmad El Sallab , Mohammed Abdou , Etienne Perot , Senthil Yogamani

This paper introduces RouteFinder, a comprehensive foundation model framework to tackle different Vehicle Routing Problem (VRP) variants. Our core idea is that a foundation model for VRPs should be able to represent variants by treating…

Many of the observations we make are biased by our decisions. For instance, the demand of items is impacted by the prices set, and online checkout choices are influenced by the assortments presented. The challenge in decision-making under…

机器学习 · 计算机科学 2025-07-02 Rares Cristian , Pavithra Harsha , Georgia Perakis , Brian Quanz

The deployment of humanoid robots in unstructured, human-centric environments requires navigation capabilities that extend beyond simple locomotion to include robust perception, provable safety, and socially aware behavior. Current…

机器人学 · 计算机科学 2025-08-12 Zifan Wang , Xun Yang , Jianzhuang Zhao , Jiaming Zhou , Teli Ma , Ziyao Gao , Arash Ajoudani , Junwei Liang

The use of electric vehicles (EV) in the last mile is appealing from both sustainability and operational cost perspectives. In addition to the inherent cost efficiency of EVs, selling energy back to the grid during peak grid demand, is a…

人工智能 · 计算机科学 2022-04-13 Ajay Narayanan , Prasant Misra , Ankush Ojha , Vivek Bandhu , Supratim Ghosh , Arunchandar Vasan

With the increasing popularity of machine learning techniques, it has become common to see prediction algorithms operating within some larger process. However, the criteria by which we train these algorithms often differ from the ultimate…

机器学习 · 计算机科学 2019-04-26 Priya L. Donti , Brandon Amos , J. Zico Kolter

We present a control approach for autonomous vehicles based on deep reinforcement learning. A neural network agent is trained to map its estimated state to acceleration and steering commands given the objective of reaching a specific target…

机器人学 · 计算机科学 2020-03-16 Andreas Folkers , Matthias Rick , Christof Büskens

Reinforcement learning has recently shown promise in learning quality solutions in many combinatorial optimization problems. In particular, the attention-based encoder-decoder models show high effectiveness on various routing problems,…

最优化与控制 · 数学 2022-12-06 Aigerim Bogyrbayeva , Taehyun Yoon , Hanbum Ko , Sungbin Lim , Hyokun Yun , Changhyun Kwon

In this work, we study vision-based end-to-end reinforcement learning on vehicle control problems, such as lane following and collision avoidance. Our controller policy is able to control a small-scale robot to follow the right-hand lane of…

机器学习 · 计算机科学 2020-12-15 András Kalapos , Csaba Gór , Róbert Moni , István Harmati

End-to-end deep reinforcement learning (DRL) for quadrotor control promises many benefits -- easy deployment, task generalization and real-time execution capability. Prior end-to-end DRL-based methods have showcased the ability to deploy…

机器人学 · 计算机科学 2024-05-07 Zhehui Huang , Zhaojing Yang , Rahul Krupani , Baskın Şenbaşlar , Sumeet Batra , Gaurav S. Sukhatme

This paper addresses the challenges of decision-making for autonomous vehicles under faults during a transport mission. A real-time decision-making problem of vehicle routing planning considering maintenance management is formulated as an…

系统与控制 · 电气工程与系统科学 2022-02-10 Xin Tao , Zhao Yuan

The recently presented idea to learn heuristics for combinatorial optimization problems is promising as it can save costly development. However, to push this idea towards practical implementation, we need better models and better ways of…

机器学习 · 统计学 2019-02-08 Wouter Kool , Herke van Hoof , Max Welling

Obstacle avoidance for unmanned aerial vehicles like quadrotors is a popular research topic. Most existing research focuses only on static environments, and obstacle avoidance in environments with multiple dynamic obstacles remains…

机器人学 · 计算机科学 2025-03-19 Xiyu Fan , Minghao Lu , Bowen Xu , Peng Lu

Mobile edge computing (MEC) is essential for next-generation mobile network applications that prioritize various performance metrics, including delays and energy consumption. However, conventional single-objective scheduling solutions…

网络与互联网体系结构 · 计算机科学 2023-07-28 Ning Yang , Junrui Wen , Meng Zhang , Ming Tang

This paper addresses the cooperative Multi-Vehicle Dynamic Pickup and Delivery Problem with Stochastic Requests (MVDPDPSR) and proposes an end-to-end centralized decision-making framework based on sequence-to-sequence, named Multi-Agent…

机器学习 · 计算机科学 2025-12-18 Zengyu Zou , Jingyuan Wang , Yixuan Huang , Junjie Wu