English
Related papers

Related papers: The Gittins Policy in the M/G/1 Queue

200 papers

We study the stochastic Multi-Armed Bandit (MAB) problem under worst-case regret and heavy-tailed reward distribution. We modify the minimax policy MOSS for the sub-Gaussian reward distribution by using saturated empirical mean to design a…

Machine Learning · Statistics 2020-11-19 Lai Wei , Vaibhav Srivastava

In this paper, we consider the average age minimization problem where a central entity schedules M users among the N available users for transmission over unreliable channels. It is well-known that obtaining the optimal policy, in this…

Information Theory · Computer Science 2020-01-14 Ali Maatouk , Saad Kriouile , Mohamad Assaad , Anthony Ephremides

We study the multi-armed bandit (MAB) problem where the agent receives a vectorial feedback that encodes many possibly competing objectives to be optimized. The goal of the agent is to find a policy, which can optimize these objectives…

Machine Learning · Computer Science 2017-06-16 Robert Busa-Fekete , Balazs Szorenyi , Paul Weng , Shie Mannor

Although many algorithms for the multi-armed bandit problem are well-understood theoretically, empirical confirmation of their effectiveness is generally scarce. This paper presents a thorough empirical study of the most popular multi-armed…

Artificial Intelligence · Computer Science 2014-02-26 Volodymyr Kuleshov , Doina Precup

As the rapidly developments of artificial intelligence and machine learning, behavior tree design in multiagent system or AI game become more important. The behavior tree design problem is highly related to the source coding in information…

Information Theory · Computer Science 2024-05-28 Zhefan Li , Pingyi Fan

We present a two-armed bandit model of decision making under uncertainty where the expected return to investing in the "risky arm" increases when choosing that arm and decreases when choosing the "safe" arm. These dynamics are natural in…

Optimization and Control · Mathematics 2017-03-22 Roland Fryer , Philipp Harms

We study the accumulation of resources within a target due to the interplay between continual delivery, driven by 1D stochastic search processes, and sequential consumption. The assumption of sequential consumption is key because it changes…

Statistical Mechanics · Physics 2025-08-27 José Giral-Barajas , Paul C Bressloff

In this paper, we consider a queueing system with multiple channels (or servers) and multiple classes of users. We aim at allocating the available channels among the users in such a way to minimize the expected total average queue length of…

Networking and Internet Architecture · Computer Science 2019-02-07 Saad Kriouile , Maialen Larranaga , Mohamad Assaad

In this paper, we study the problem of Gaussian process (GP) bandits under relaxed optimization criteria stating that any function value above a certain threshold is "good enough". On the theoretical side, we study various {\em lenient…

Machine Learning · Statistics 2021-05-27 Xu Cai , Selwyn Gomes , Jonathan Scarlett

We consider a time-slotted job-assignment system consisting of a central server, $N$ task-specific networks of machines, and multiple users. Each network specializes in executing a distinct type of task. Users stochastically generate jobs…

Information Theory · Computer Science 2026-02-03 Subhankar Banerjee , Sennur Ulukus

We study critical GI/G/1 queues under finite second moment assumptions. We show that the busy period distribution is regularly varying with index half. We also review previously known M/G/1/ and M/M/1 derivations, yielding exact asymptotics…

Probability · Mathematics 2022-04-12 Yoni Nazarathy , Zbigniew Palmowski

We formulate a control problem for a GI/GI/N+GI queue, whose objective is to trade off the long-run average operational costs (i.e., abandonment costs and holding costs) with server utilization costs. To solve the control problem, we…

Optimization and Control · Mathematics 2023-11-08 Yueyang Zhong , Amy R. Ward , Amber L. Puha

We introduce an original minimax framework for finite-time performance analysis in queueing control and propose a surprisingly simple Lyapunov-based scheduling policy with superior finite-time performance. The framework quantitatively…

Optimization and Control · Mathematics 2025-12-01 Yujie Liu , Vincent Y. F. Tan , Yunbei Xu

We examine a generalised queuing model which we call the G/G/n/G/+ model, which encompasses the G/G/n and G/G/n/s models as special cases. Our model accommodates useful generalisations in user behaviour and limitations on the facilities for…

Computation · Statistics 2021-11-16 Ben O'Neill

In this paper continuity theorems are established for the number of losses during a busy period of the $M/M/1/n$ queue. We consider an $M/GI/1/n$ queueing system where the service time probability distribution, slightly different in a…

Probability · Mathematics 2008-08-01 Vyacheslav M. Abramov

We consider the FCFS $GI/GI/n$ queue in the Halfin-Whitt heavy traffic regime, and prove bounds for the steady-state probability of delay (s.s.p.d.) for generally distributed processing times. We prove that there exist $\epsilon_1,…

Probability · Mathematics 2016-09-02 David A. Goldberg

We study a single server FIFO queue that offers general service. Each of n customers enter the queue at random time epochs that are inde- pendent and identically distributed. We call this the random scattering traffic model, and the…

Probability · Mathematics 2017-08-21 Peter W. Glynn , Harsha Honnappa

Recently it was shown that the response time of First-Come-First-Served (FCFS) scheduling can be stochastically and asymptotically improved upon by the {\it Nudge} scheduling algorithm in case of light-tailed job size distributions. Such…

Performance · Computer Science 2024-04-15 Nils Charlet , Benny Van Houdt

This paper introduces the first asymptotically optimal strategy for a multi armed bandit (MAB) model under side constraints. The side constraints model situations in which bandit activations are limited by the availability of certain…

Machine Learning · Statistics 2025-02-10 Apostolos N. Burnetas , Odysseas Kanavetas , Michael N. Katehakis

We study the stationary sojourn time distribution in an M/G/1 queue operating under heavy traffic. It is known that the sojourn time converges to an exponential distribution in the limit. Our focus is on obtaining pre-asymptotic,…

Probability · Mathematics 2026-01-21 Bihan Chatterjee , Siva Theja Maguluri , Debankur Mukherjee
‹ Prev 1 4 5 6 7 8 10 Next ›