中文
相关论文

相关论文: Double Q-Learning for Citizen Relocation During Na…

200 篇论文

This paper proposes a reinforcement learning approach for traffic control with the adaptive horizon. To build the controller for the traffic network, a Q-learning-based strategy that controls the green light passing time at the network…

系统与控制 · 计算机科学 2019-04-01 Wentao Chen , Tehuan Chen , Guang Lin

This study proposes two new dynamic assignment algorithms to match refugees and asylum seekers to geographic localities within a host country. The first, currently implemented in a multi-year randomized control trial in Switzerland, seeks…

最优化与控制 · 数学 2024-05-28 Kirk Bansak , Elisabeth Paulson

Learning-based congestion control (CC), including Reinforcement-Learning, promises efficient CC in a fast-changing networking landscape, where evolving communication technologies, applications and traffic workloads pose severe challenges to…

网络与互联网体系结构 · 计算机科学 2026-04-17 Mihai Mazilu , Luca Giacomoni , George Parisis

There has been increasing awareness of the difficulties in reaching and extracting people from mass casualty scenarios, such as those arising from natural disasters. While platforms have been designed to consider reaching casualties and…

机器人学 · 计算机科学 2023-09-28 Elizabeth Peiros , Zih-Yun Chiu , Yuheng Zhi , Nikhil Shinde , Michael C. Yip

Reinforcement learning algorithms have been widely used for decision-making tasks in various domains. However, the performance of these algorithms can be impacted by high variance and instability, particularly in environments with noise or…

机器学习 · 统计学 2026-03-31 Saunak Kumar Panda , Tong Li , Ruiqi Liu , Yisha Xiang

Reinforcement learning (RL) algorithms have become indispensable tools in artificial intelligence, empowering agents to acquire optimal decision-making policies through interactions with their environment and feedback mechanisms. This study…

机器学习 · 计算机科学 2024-03-28 Ergon Cugler de Moraes Silva

This paper designs a sequential repeated game of a micro-founded society with three types of agents: individuals, insurers, and a government. Nascent to economics literature, we use Reinforcement Learning (RL), closely related to…

多智能体系统 · 计算机科学 2022-07-05 Menna Hassan , Nourhan Sakr , Arthur Charpentier

This tutorial contrasts probabilistic modeling and robust optimization to determine decisions in humanitarian logistics, specifically supply chains subject to adversarial (natural and human) disruptions. Natural disruptions induce dispatch…

最优化与控制 · 数学 2026-04-06 Justin Kilb , Daniel Bienstock , Alexandra M. Newman

Autonomous agents often require multiple strategies to solve complex tasks, but determining when to switch between strategies remains challenging. This research introduces a reinforcement learning technique to learn switching thresholds…

机器学习 · 计算机科学 2025-12-09 Chris Tava

A long-term goal of reinforcement learning is to design agents that can autonomously interact and learn in the world. A critical challenge to such autonomy is the presence of irreversible states which require external assistance to recover…

机器学习 · 计算机科学 2022-10-20 Annie Xie , Fahim Tajwar , Archit Sharma , Chelsea Finn

We present a simple yet efficient Hybrid Classifier based on Deep Learning and Reinforcement Learning. Q-Learning is used with two Q-states and four actions. Conventional techniques use feature maps extracted from Convolutional Neural…

计算机视觉与模式识别 · 计算机科学 2020-08-11 Abdul Mueed Hafiz , Ghulam Mohiuddin Bhat

Path planning for mobile robots in large dynamic environments is a challenging problem, as the robots are required to efficiently reach their given goals while simultaneously avoiding potential conflicts with other robots or dynamic…

机器人学 · 计算机科学 2020-09-15 Binyu Wang , Zhe Liu , Qingbiao Li , Amanda Prorok

Spatial disorientation is a leading cause of fatal aircraft accidents. This paper explores the potential of AI agents to aid pilots in maintaining balance and preventing unrecoverable losses of control by offering cues and corrective…

Robot-assisted navigation is a perfect example of a class of applications requiring flexible control approaches. When the human is reliable, the robot should concede space to their initiative. When the human makes inappropriate choices the…

机器人学 · 计算机科学 2023-12-25 Placido Falqueto , Alessandro Antonucci , Luigi Palopoli , Daniele Fontanelli

Clinical decision-making often involves selecting tests that are costly, invasive, or time-consuming, motivating individualized, sequential strategies for what to measure and when to stop ascertaining. We study the problem of learning…

机器学习 · 统计学 2026-04-16 Doudou Zhou , Yiran Zhang , Dian Jin , Yingye Zheng , Lu Tian , Tianxi Cai

A central goal in ecology is to understand how biodiversity is maintained. Previous theoretical works have employed the rock-paper-scissors (RPS) game as a toy model, demonstrating that population mobility is crucial in determining the…

种群与进化 · 定量生物学 2026-05-21 Kaiwen Jiang , Chenyang Zhao , Shengfeng Deng , Weiran Cai , Jiqiang Zhang , Li Chen

This work addresses the coordination problem of multiple robots with the goal of finding specific hazardous targets in an unknown area and dealing with them cooperatively. The desired behaviour for the robotic system entails multiple…

机器人学 · 计算机科学 2019-03-29 Nunzia Palmieri , Xin-She Yang , Floriano De Rango , Amilcare Francesco Santamaria

With the increasing frequency of major natural disasters, understanding their political consequences is of paramount importance for democratic accountability. The existing literature is deeply divided, with some studies finding that voters…

综合经济学 · 经济学 2025-07-22 Nima Taheri Hosseinkhani

This paper investigates the automatic exploration problem under the unknown environment, which is the key point of applying the robotic system to some social tasks. The solution to this problem via stacking decision rules is impossible to…

机器人学 · 计算机科学 2020-07-24 Haoran Li , Qichao Zhang , Dongbin Zhao

Most learning algorithms with formal regret guarantees assume that all mistakes are recoverable and essentially rely on trying all possible behaviors. This approach is problematic when some mistakes are "catastrophic", i.e., irreparable. We…

机器学习 · 计算机科学 2025-08-07 Benjamin Plaut , Hanlin Zhu , Stuart Russell