English
Related papers

Related papers: An Enhanced Search Technique for Managing Partial …

200 papers

A dynamic treatment regimen (DTR) is a set of decision rules to personalize treatments for an individual using their medical history. The Q-learning-based Q-shared algorithm has been used to develop DTRs that involve decision rules shared…

Machine Learning · Statistics 2024-12-05 Palash Ghosh , Xinru Wang , Trikay Nalamada , Shruti Agarwal , Maria Jahja , Bibhas Chakraborty

This paper proposes a reinforcement learning approach for traffic control with the adaptive horizon. To build the controller for the traffic network, a Q-learning-based strategy that controls the green light passing time at the network…

Systems and Control · Computer Science 2019-04-01 Wentao Chen , Tehuan Chen , Guang Lin

Peer-to-peer (P2P) networks establish loosely coupled application-level overlays on top of the Internet to facilitate efficient sharing of resources. It can be roughly classified as either structured or unstructured networks. Without…

Networking and Internet Architecture · Computer Science 2010-06-24 R. Anusuya , V. Kavitha , E. Golden Julie

Sequential decision tasks with incomplete information are characterized by the exploration problem; namely the trade-off between further exploration for learning more about the environment and immediate exploitation of the accrued…

Artificial Intelligence · Computer Science 2013-02-21 Grigoris I. Karakoulas

We study the problem of representation transfer in offline Reinforcement Learning (RL), where a learner has access to episodic data from a number of source tasks collected a priori, and aims to learn a shared representation to be used in…

Machine Learning · Computer Science 2024-02-21 Avinandan Bose , Simon Shaolei Du , Maryam Fazel

The drift-plus-penalty method is a Lyapunov optimisation technique commonly applied to network routing problems. It reduces the original stochastic planning task to a sequence of greedy optimizations, enabling the design of distributed…

Systems and Control · Electrical Eng. & Systems 2025-09-12 Ahmed Rashwan , Keith Briggs , Chris Budd

In this paper, we study a variant of the dynamic ridesharing problem with a specific focus on peak hours: Given a set of drivers and rider requests, we aim to match drivers to each rider request by achieving two objectives: maximizing the…

Databases · Computer Science 2020-04-07 Hui Luo , Zhifeng Bao , Farhana M. Choudhury , J. Shane Culpepper

Reinforcement Learning algorithms have recently been proposed to learn time-sequential control policies in the field of autonomous driving. Direct applications of Reinforcement Learning algorithms with discrete action space will yield…

Machine Learning · Computer Science 2019-12-03 Pin Wang , Hanhan Li , Ching-Yao Chan

Path Planning methods for autonomous control of Unmanned Aerial Vehicle (UAV) swarms are on the rise because of all the advantages they bring. There are more and more scenarios where autonomous control of multiple UAVs is required. Most of…

Artificial Intelligence · Computer Science 2023-08-28 Alejandro Puente-Castro , Daniel Rivero , Eurico Pedrosa , Artur Pereira , Nuno Lau , Enrique Fernandez-Blanco

This paper presents a deep Q-network (DQN)-based gain-scheduling framework for safety-critical quadcopter trajectory tracking. Instead of directly learning control inputs, the proposed approach selects from a finite set of pre-certified…

Systems and Control · Electrical Eng. & Systems 2026-03-04 Hossein Rastgoftar , Muhammad J. H. Zahed

Efficient retrieval of information is of key importance when using Big Data systems. In large scale-out data architectures, data are distributed and replicated across several machines. Queries/tasks to such data architectures, are sent to a…

Q-learning is widely used to optimize wireless networks with unknown system dynamics. Recent advancements include ensemble multi-environment hybrid Q-learning algorithms, which utilize multiple Q-learning algorithms across structurally…

Signal Processing · Electrical Eng. & Systems 2024-09-02 Talha Bozkus , Urbashi Mitra

A model of providing service in a P2P network is analyzed. It is shown that by adding a scrip system, a mechanism that admits a reasonable Nash equilibrium that reduces free riding can be obtained. The effect of varying the total amount of…

Computer Science and Game Theory · Computer Science 2007-05-30 Eric j. Friedman , Joseph Y. Halpern , Ian Kash

Reinforcement Learning is gaining attention by the wireless networking community due to its potential to learn good-performing configurations only from the observed results. In this work we propose a stateless variation of Q-learning, which…

Networking and Internet Architecture · Computer Science 2017-08-30 Francesc Wilhelmi , Boris Bellalta , Cristina Cano , Anders Jonsson

Monitoring and patrolling large water resources is a major challenge for conservation. The problem of acquiring data of an underlying environment that usually changes within time involves a proper formulation of the information. The use of…

Robotics · Computer Science 2022-10-18 Samuel Yanes Luis , Daniel Gutiérrez Reina , Sergio Toral Marín

Spurred by the growth of transportation network companies and increasing data capabilities, vehicle routing and ride-matching algorithms can improve the efficiency of private transportation services. However, existing routing solutions do…

Systems and Control · Computer Science 2018-10-25 Ian Schneider , Jun Jie Joseph Kuan , Mardavij Roozbehani , Munther Dahleh

In this paper, the trajectory optimization problem for a multi-aerial base station (ABS) communication network is investigated. The objective is to find the trajectory of the ABSs so that the sum-rate of the users served by each ABS is…

Signal Processing · Electrical Eng. & Systems 2019-07-02 Behzad Khamidehi , Elvino S. Sousa

The use of target networks in deep reinforcement learning is a widely popular solution to mitigate the brittleness of semi-gradient approaches and stabilize learning. However, target networks notoriously require additional memory and delay…

Machine Learning · Computer Science 2026-03-02 Théo Vincent , Yogesh Tripathi , Tim Faust , Abdullah Akgül , Yaniv Oren , Melih Kandemir , Jan Peters , Carlo D'Eramo

Increasing demand for Large Language Models (LLMs) services imposes substantial deployment and computation costs on providers. LLM routing offers a cost-efficient solution by directing queries to the optimal LLM based on model and query…

Databases · Computer Science 2025-12-16 Fangzhou Wu , Sandeep Silwal

Bike-sharing systems play a crucial role in easing traffic congestion and promoting healthier lifestyles. However, ensuring their reliability and user acceptance requires effective strategies for rebalancing bikes. This study introduces a…

Machine Learning · Computer Science 2024-06-04 Jiaqi Liang , Defeng Liu , Sanjay Dominik Jena , Andrea Lodi , Thibaut Vidal