中文
相关论文

相关论文: Zero-shot Transfer Learning of Driving Policy via …

200 篇论文

In this article, we demonstrate a zero-shot transfer of an autonomous driving policy from simulation to University of Delaware's scaled smart city with adversarial multi-agent reinforcement learning, in which an adversary attempts to…

Semi-cooperative behaviors are intrinsic properties of human drivers and should be considered for autonomous driving. In addition, new autonomous planners can consider the social value orientation (SVO) of human drivers to generate…

计算与语言 · 计算机科学 2023-10-02 Noam Buckman , Sertac Karaman , Daniela Rus

Using deep reinforcement learning, we train control policies for autonomous vehicles leading a platoon of vehicles onto a roundabout. Using Flow, a library for deep reinforcement learning in micro-simulators, we train two policies, one…

系统与控制 · 计算机科学 2019-02-26 Kathy Jang , Eugene Vinitsky , Behdad Chalaki , Ben Remer , Logan Beaver , Andreas Malikopoulos , Alexandre Bayen

This paper studies a stochastic dynamic game between two competing teams, each consisting of a network of collaborating agents. Unlike fully cooperative settings, where all agents share a common objective, each team in this game aims to…

多智能体系统 · 计算机科学 2025-04-29 Yike Zhao , Haoyuan Cai , Ali H. Sayed

We report on a study that employs an in-house developed simulation infrastructure to accomplish zero shot policy transferability for a control policy associated with a scale autonomous vehicle. We focus on implementing policies that require…

We use model-free reinforcement learning, extensive simulation, and transfer learning to develop a continuous control algorithm that has good zero-shot performance in a real physical environment. We train a simulated agent to act optimally…

人工智能 · 计算机科学 2018-03-09 M Ferguson , K. H. Law

Accurate trajectory prediction of surrounding vehicles (SVs) is crucial for autonomous driving systems to avoid misguided decisions and potential accidents. However, achieving reliable predictions in highly dynamic and complex traffic…

机器人学 · 计算机科学 2025-09-23 Xiao Zhou , Zengqi Peng , Jun Ma

Traffic simulation, complementing real-world data with a long-tail distribution, allows for effective evaluation and enhancement of the ability of autonomous vehicles to handle accident-prone scenarios. Simulating such safety-critical…

机器人学 · 计算机科学 2025-03-10 Zherui Huang , Xing Gao , Guanjie Zheng , Licheng Wen , Xuemeng Yang , Xiao Sun

The scenario-based testing of operational vehicle safety presents a set of principal other vehicle (POV) trajectories that seek to force the subject vehicle (SV) into a certain safety-critical situation. Current scenarios are mostly (i)…

机器人学 · 计算机科学 2021-05-24 Linda Capito , Bowen Weng , Umit Ozguner , Keith Redmill

Multi-Agent Reinforcement Learning (MARL) has become a promising solution for constructing a multi-agent autonomous driving system (MADS) in complex and dense scenarios. But most methods consider agents acting selfishly, which leads to…

机器人学 · 计算机科学 2023-10-06 Jintao Xue , Dongkun Zhang , Rong Xiong , Yue Wang , Eryun Liu

With the commercial application of automated vehicles (AVs), the sharing of roads between AVs and human-driven vehicles (HVs) becomes a common occurrence in the future. While research has focused on improving the safety and reliability of…

机器人学 · 计算机科学 2023-07-03 Yan Tong , Licheng Wen , Pinlong Cai , Daocheng Fu , Song Mao , Yikang Li

Robotic agents must adopt existing social conventions in order to be effective teammates. These social conventions, such as driving on the right or left side of the road, are arbitrary choices among optimal policies, but all agents on a…

人工智能 · 计算机科学 2020-10-09 Mycal Tucker , Yilun Zhou , Julie Shah

The integration of Autonomous Vehicles (AVs) into existing human-driven traffic systems poses considerable challenges, especially within environments where human and machine interactions are frequent and complex, such as at unsignalized…

机器人学 · 计算机科学 2024-04-05 Jiaqi Liu , Xiao Qi , Peng Hang , Jian Sun

In modern transportation networks, adversaries can manipulate routing algorithms using false data injection attacks, such as simulating heavy traffic with multiple devices running crowdsourced navigation applications, to mislead vehicles…

人工智能 · 计算机科学 2026-03-13 Taha Eghtesad , Yevgeniy Vorobeychik , Aron Laszka

Deploying autonomous driving systems requires robustness against long-tail scenarios that are rare but safety-critical. While adversarial training offers a promising solution, existing methods typically decouple scenario generation from…

机器学习 · 计算机科学 2026-03-17 Tong Nie , Yihong Tang , Junlin He , Yuewen Mei , Jie Sun , Lijun Sun , Wei Ma , Jian Sun

Most of the current studies on autonomous vehicle decision-making and control tasks based on reinforcement learning are conducted in simulated environments. The training and testing of these studies are carried out under rule-based…

系统与控制 · 电气工程与系统科学 2024-04-22 Yuan Lin , Antai Xie , Xiao Liu

Active traffic management with autonomous vehicles offers the potential for reduced congestion and improved traffic flow. However, developing effective algorithms for real-world scenarios requires overcoming challenges related to…

机器学习 · 计算机科学 2024-09-04 Shengchao Yan , Lukas König , Wolfram Burgard

In social psychology, Social Value Orientation (SVO) describes an individual's propensity to allocate resources between themself and others. In reinforcement learning, SVO has been instantiated as an intrinsic motivation that remaps an…

Transportation and traffic are currently undergoing a rapid increase in terms of both scale and complexity. At the same time, an increasing share of traffic participants are being transformed into agents driven or supported by artificial…

机器学习 · 计算机科学 2018-10-24 Mark Schutera , Niklas Goby , Dirk Neumann , Markus Reischl

Off-policy learning methods seek to derive an optimal policy directly from a fixed dataset of prior interactions. This objective presents significant challenges, primarily due to the inherent distributional shift and value function…

机器学习 · 计算机科学 2026-02-03 Arip Asadulaev , Maksim Bobrin , Salem Lahlou , Dmitry Dylov , Fakhri Karray , Martin Takac
‹ 上一页 1 2 3 10 下一页 ›