中文
相关论文

相关论文: DQN Control Solution for KDD Cup 2021 City Brain C…

200 篇论文

As reinforcement learning (RL) scales to solve increasingly complex tasks, interest continues to grow in the fields of AI safety and machine ethics. As a contribution to these fields, this paper introduces an extension to Deep Q-Networks…

机器学习 · 计算机科学 2019-06-27 Bart Bussmann , Jacqueline Heinerman , Joel Lehman

Discretionary lane-change is one of the critical challenges for autonomous vehicle (AV) design due to its significant impact on traffic efficiency. Existing intelligent lane-change solutions have primarily focused on optimizing the…

计算机与社会 · 计算机科学 2023-03-17 Lokesh Chandra Das , Myounggyu Won

Opponent modeling is necessary in multi-agent settings where secondary agents with competing goals also adapt their strategies, yet it remains challenging because strategies interact with each other and change. Most previous work focuses on…

机器学习 · 计算机科学 2016-09-20 He He , Jordan Boyd-Graber , Kevin Kwok , Hal Daumé

Spectrum sharing among users is a fundamental problem in the management of any wireless network. In this paper, we discuss the problem of distributed spectrum collaboration without central management under general unknown channels. Since…

信号处理 · 电气工程与系统科学 2021-04-07 Pranav M. Pawar , Amir Leshem

High-radix interconnects such as Dragonfly and its variants rely on adaptive routing to balance network traffic for optimum performance. Ideally, adaptive routing attempts to forward packets between minimal and non-minimal paths with the…

网络与互联网体系结构 · 计算机科学 2024-04-05 Yao Kang , Xin Wang , Zhiling Lan

Q-learning is a widely used reinforcement learning technique for solving path planning problems. It primarily involves the interaction between an agent and its environment, enabling the agent to learn an optimal strategy that maximizes…

机器人学 · 计算机科学 2024-12-18 Yiming Ji , Kaijie Yun , Yang Liu , Zongwu Xie , Hong Liu

This paper proposes a reinforcement learning approach for traffic control with the adaptive horizon. To build the controller for the traffic network, a Q-learning-based strategy that controls the green light passing time at the network…

系统与控制 · 计算机科学 2019-04-01 Wentao Chen , Tehuan Chen , Guang Lin

Reinforcement learning (RL) is attracting attention as an effective way to solve sequential optimization problems that involve high dimensional state/action space and stochastic uncertainties. Many such problems involve constraints…

机器学习 · 计算机科学 2021-04-01 Haeun Yoo , Victor M. Zavala , Jay H. Lee

With the advent of ride-sharing services, there is a huge increase in the number of people who rely on them for various needs. Most of the earlier approaches tackling this issue required handcrafted functions for estimating travel times and…

机器学习 · 计算机科学 2020-06-22 Oscar de Lima , Hansal Shah , Ting-Sheng Chu , Brian Fogelson

In this paper, we investigate traffic signal control in a network of interconnected intersections, aiming to balance lane-level vehicle densities through optimal green-time allocation. We develop a two-lane traffic flow model that…

系统与控制 · 电气工程与系统科学 2025-12-09 Xinfeng Ru , Ting Bai , Weiguo Xia , Andreas A. Malikopoulos

We propose a distributed algorithm for controlling traffic signals, allowing constraints such as periodic switching sequences of phases and minimum and maximum green time to be incorporated. Our algorithm is adapted from backpressure…

系统与控制 · 计算机科学 2014-07-07 Tichakorn Wongpiromsarn , Tawit Uthaicharoenpong , Emilio Frazzoli , Yu Wang , Danwei Wang

Connected Autonomous Vehicles will make autonomous intersection management a reality replacing traditional traffic signal control. Autonomous intersection management requires time and speed adjustment of vehicles arriving at an intersection…

多智能体系统 · 计算机科学 2022-02-10 Udesh Gunarathna , Shanika Karunasekara , Renata Borovica-Gajic , Egemen Tanin

Multi-agent Deep Reinforcement Learning (MADRL) based traffic signal control becomes a popular research topic in recent years. To alleviate the scalability issue of completely centralized RL techniques and the non-stationarity issue of…

人工智能 · 计算机科学 2023-09-08 Hankang Gu , Shangbo Wang , Xiaoguang Ma , Dongyao Jia , Guoqiang Mao , Eng Gee Lim , Cheuk Pong Ryan Wong

Inspired by Double Q-learning algorithm, the Double-DQN (DDQN) algorithm was originally proposed in order to address the overestimation issue in the original DQN algorithm. The DDQN has successfully shown both theoretically and empirically…

人工智能 · 计算机科学 2024-10-30 Shervin Halat , Mohammad Mehdi Ebadzadeh , Kiana Amani

We present a distributed quasi-Newton (DQN) method, which enables a group of agents to compute an optimal solution of a separable multi-agent optimization problem locally using an approximation of the curvature of the aggregate objective…

最优化与控制 · 数学 2024-09-30 Ola Shorinwa , Mac Schwager

The recent advancements in cloud services, Internet of Things (IoT) and Cellular networks have made cloud computing an attractive option for intelligent traffic signal control (ITSC). Such a method significantly reduces the cost of cables,…

信号处理 · 电气工程与系统科学 2020-03-09 Rusheng Zhang , Xinze Zhou , Ozan K. Tonguz

Deep reinforcement learning for high dimensional, hierarchical control tasks usually requires the use of complex neural networks as functional approximators, which can lead to inefficiency, instability and even divergence in the training…

机器学习 · 计算机科学 2019-11-26 Yuguang Yang

Autonomous driving has been at the forefront of public interest, and a pivotal debate to widespread concerns is safety in the transportation system. Deep reinforcement learning (DRL) has been applied to autonomous driving to provide…

人工智能 · 计算机科学 2022-01-21 Zehong Cao , Jie Yun

Deep reinforcement learning (RL) has achieved several high profile successes in difficult decision-making problems. However, these algorithms typically require a huge amount of data before they reach reasonable performance. In fact, their…

This paper introduces the QDQN-DPER framework to enhance the efficiency of quantum reinforcement learning (QRL) in solving sequential decision tasks. The framework incorporates prioritized experience replay and asynchronous training into…

量子物理 · 物理学 2023-04-20 Samuel Yen-Chi Chen
‹ 上一页 1 8 9 10 下一页 ›