English
Related papers

Related papers: Double Q-Learning for Citizen Relocation During Na…

200 papers

Q-learning is a stochastic approximation version of the classic value iteration. The literature has established that Q-learning suffers from both maximization bias and slower convergence. Recently, multi-step algorithms have shown practical…

Machine Learning · Computer Science 2024-07-03 Antony Vijesh , Shreyas S R

Development of navigation algorithms is essential for the successful deployment of robots in rapidly changing hazardous environments for which prior knowledge of configuration is often limited or unavailable. Use of traditional…

Robotics · Computer Science 2022-11-11 Paul Blum , Peter Crowley , George Lykotrafitis

The open streets initiative "opens" streets to pedestrians and bicyclists by closing them to cars and trucks. The initiative, adopted by many cities across North America, increases community space in urban environments. But could open…

Machine Learning · Computer Science 2023-12-14 R. Teal Witter , Lucas Rosenblatt

State-of-the-art emergency navigation approaches are designed to evacuate civilians during a disaster based on real-time decisions using a pre-defined algorithm and live sensory data. Hence, casualties caused by the poor decisions and…

Other Computer Science · Computer Science 2015-01-06 Huibo Bi , Erol Gelenbe

The optimistic nature of the Q-learning target leads to an overestimation bias, which is an inherent problem associated with standard $Q-$learning. Such a bias fails to account for the possibility of low returns, particularly in risky…

Machine Learning · Computer Science 2021-11-05 Thommen George Karimpanal , Hung Le , Majid Abdolshah , Santu Rana , Sunil Gupta , Truyen Tran , Svetha Venkatesh

The dynamic vehicle dispatching problem corresponds to deciding which vehicles to assign to requests that arise stochastically over time and space. It emerges in diverse areas, such as in the assignment of trucks to loads to be transported;…

Artificial Intelligence · Computer Science 2023-07-17 Edyvalberty Alenquer Cordeiro , Anselmo Ramalho Pitombeira-Neto

Decision-making is critical for lane change in autonomous driving. Reinforcement learning (RL) algorithms aim to identify the values of behaviors in various situations and thus they become a promising pathway to address the decision-making…

Robotics · Computer Science 2022-07-08 Jingda Wu , Wenhui Huang , Niels de Boer , Yanghui Mo , Xiangkun He , Chen Lv

Reinforcement Learning has applications in field of mechatronics, robotics, and other resource-constrained control system. Problem of resource allocation is primarily solved using traditional predefined techniques and modern deep learning…

Machine Learning · Computer Science 2021-06-18 Neel Gandhi , Shakti Mishra

Emergency evacuation describes a complex situation involving time-critical decision-making by evacuees. Mobile robots are being actively explored as a potential solution to provide timely guidance. In this work, we study a robot-guided…

Robotics · Computer Science 2024-01-15 Tongjia Zheng , Zhenyuan Yuan , Mollik Nayyar , Alan R. Wagner , Minghui Zhu , Hai Lin

A wide range of sustainability and grid-integration strategies depend on workload shifting, which aligns the timing of energy consumption with external signals such as grid curtailment events, carbon intensity, or time-of-use electricity…

Data Structures and Algorithms · Computer Science 2025-10-01 Ezra Johnson , Adam Lechowicz , Mohammad Hajiesmaili

In natural ecosystems and human societies, self-organized resource allocation and policy synergy are ubiquitous and significant. This work focuses on the synergy between Dual Reinforcement Learning Policies in the Minority Game (DRLP-MG) to…

Adaptation and Self-Organizing Systems · Physics 2025-09-29 Zhen-Na Zhang , Guo-Zhong Zhen , Li Chen , Chao-Ran Cai , Sheng-Feng Deng , Bin-Quan Li , Ji-Qiang Zhang

Path Planning methods for autonomous control of Unmanned Aerial Vehicle (UAV) swarms are on the rise because of all the advantages they bring. There are more and more scenarios where autonomous control of multiple UAVs is required. Most of…

Artificial Intelligence · Computer Science 2023-08-28 Alejandro Puente-Castro , Daniel Rivero , Eurico Pedrosa , Artur Pereira , Nuno Lau , Enrique Fernandez-Blanco

Bike-sharing systems (BSS) provide a sustainable urban mobility solution, but ensuring their reliability requires effective rebalancing strategies to address stochastic demand and prevent station imbalances. This paper proposes…

Machine Learning · Computer Science 2025-11-27 Jiaqi Liang , Defeng Liu , Sanjay Dominik Jena , Andrea Lodi , Thibaut Vidal

Managing the response to natural disasters effectively can considerably mitigate their devastating impact. This work explores the potential of using supervised hybrid quantum machine learning to optimize emergency evacuation plans for cars…

Social robot navigation is an evolving research field that aims to find efficient strategies to safely navigate dynamic environments populated by humans. A critical challenge in this domain is the accurate modeling of human motion, which…

Human-Computer Interaction · Computer Science 2025-07-01 Tommaso Van Der Meer , Andrea Garulli , Antonio Giannitrapani , Renato Quartullo

In a Human-in-the-Loop paradigm, a robotic agent is able to act mostly autonomously in solving a task, but can request help from an external expert when needed. However, knowing when to request such assistance is critical: too few requests…

Evacuee routing algorithms in emergency typically adopt one single criterion to compute desired paths and ignore the specific requirements of users caused by different physical strength, mobility and level of resistance to hazard. In this…

Other Computer Science · Computer Science 2015-01-23 Olumide J. Akinwande , Huibo Bi

Cooperation emergence in multi-agent systems represents a fundamental statistical physics problem where microscopic learning rules drive macroscopic collective behavior transitions. We propose a Q-learning-based variant of adaptive rewiring…

Physics and Society · Physics 2025-09-04 Yi-Ning Weng , Hsuan-Wei Lee

Sequential decision tasks with incomplete information are characterized by the exploration problem; namely the trade-off between further exploration for learning more about the environment and immediate exploitation of the accrued…

Artificial Intelligence · Computer Science 2013-02-21 Grigoris I. Karakoulas

Floods are one of nature's most catastrophic calamities which cause irreversible and immense damage to human life, agriculture, infrastructure and socio-economic system. Several studies on flood catastrophe management and flood forecasting…