中文
相关论文

相关论文: The Gittins Policy in the M/G/1 Queue

200 篇论文

We study a generalization of the $M/G/1$ system (denoted by $rM/G/1$) with independent and identically distributed (iid) service times and with an arrival process whose arrival rate $\lambda_0f(r)$ depends on the remaining service time $r$…

概率论 · 数学 2017-10-05 Benjamin Legros , Ali Devin Sezer

We consider a system consisting of a single transmitter/receiver pair and $N$ channels over which they may communicate. Packets randomly arrive to the transmitter's queue and wait to be successfully sent to the receiver. The transmitter may…

性能 · 计算机科学 2020-05-15 Thomas Stahlbuhk , Brooke Shrader , Eytan Modiano

We consider a novel stochastic multi-armed bandit setting, where playing an arm makes it unavailable for a fixed number of time slots thereafter. This models situations where reusing an arm too often is undesirable (e.g. making the same…

机器学习 · 计算机科学 2024-07-31 Soumya Basu , Rajat Sen , Sujay Sanghavi , Sanjay Shakkottai

We study the generalized linear contextual bandit problem within the constraints of limited adaptivity. In this paper, we present two algorithms, $\texttt{B-GLinCB}$ and $\texttt{RS-GLinCB}$, that address, respectively, two prevalent…

机器学习 · 计算机科学 2025-10-29 Ayush Sawarni , Nirjhar Das , Siddharth Barman , Gaurav Sinha

We study a sampling and transmission scheduling problem for multi-source remote estimation, where a scheduler determines when to take samples from multiple continuous-time Gauss-Markov processes and send the samples over multiple channels…

网络与互联网体系结构 · 计算机科学 2024-04-17 Tasmeen Zaman Ornee , Yin Sun

We consider the problem of revenue-optimal dynamic mechanism design in settings where agents' types evolve over time as a function of their (both public and private) experience with items that are auctioned repeatedly over an infinite…

计算机科学与博弈论 · 计算机科学 2010-10-18 Sham M. Kakade , Ilan Lobel , Hamid Nazerzadeh

We present a formal model of human decision-making in explore-exploit tasks using the context of multi-armed bandit problems, where the decision-maker must choose among multiple options with uncertain rewards. We address the standard…

机器学习 · 计算机科学 2019-12-23 Paul Reverdy , Vaibhav Srivastava , Naomi E. Leonard

We compute the stationary performance metrics of a single server $M^X/G/1$ queue under a class of generalized processor-sharing scheduling policies that are proposed by Grishechkin. This class of processor-sharing policies allow service…

概率论 · 数学 2021-03-18 Yingdong Lu

The paper develops a novel motion model, called Generalized Multi-Speed Dubins Motion Model (GMDM), which extends the Dubins model by considering multiple speeds. While the Dubins model produces time-optimal paths under a constant speed…

机器人学 · 计算机科学 2025-04-14 James P. Wilson , Shalabh Gupta , Thomas A. Wettergren

The single server queue with multiple customer types and semi-Markovian service times, sometimes referred to as the $M/SM/1$ queue, has been well-studied since its introduction by Neuts in 1966. In this paper, we apply an extension of this…

概率论 · 数学 2018-12-07 Abhishek , Marko Boon , Rudesindo Núñez-Queija

We consider a single-hop switched queueing network. Amongst a plethora of applications, these networks have been used to model wireless networks and input queued switches. The MaxWeight scheduling policies have proved popular, chiefly,…

最优化与控制 · 数学 2013-01-17 N. S. Walton

We study the multi-armed bandit problem with arms which are Markov chains with rewards. In the finite-horizon setting, the celebrated Gittins indices do not apply, and the exact solution is intractable. We provide approximation algorithms…

数据结构与算法 · 计算机科学 2016-09-14 Will Ma

This paper is about index policies for minimizing (frequentist) regret in a stochastic multi-armed bandit model, inspired by a Bayesian view on the problem. Our main contribution is to prove that the Bayes-UCB algorithm, which relies on…

机器学习 · 统计学 2017-11-07 Emilie Kaufmann

For distributed computing environment, we consider the empirical risk minimization problem and propose a distributed and communication-efficient Newton-type optimization method. At every iteration, each worker locally finds an Approximate…

机器学习 · 计算机科学 2018-09-12 Shusen Wang , Farbod Roosta-Khorasani , Peng Xu , Michael W. Mahoney

We consider a general queueing system with price-sensitive customers in which the service provider seeks to balance two objectives, maximizing the average revenue rate and minimizing the average queue length. Customers arrive according to a…

数据结构与算法 · 计算机科学 2025-12-09 Jacob Bergquist , Adam N. Elmachtoub

This paper considers what we propose to call multi-gear bandits, which are Markov decision processes modeling a generic dynamic and stochastic project fueled by a single resource and which admit multiple actions representing gears of…

最优化与控制 · 数学 2026-01-21 José Niño-Mora

Robots often rely on a repertoire of previously-learned motion policies for performing tasks of diverse complexities. When facing unseen task conditions or when new task requirements arise, robots must adapt their motion policies…

机器学习 · 计算机科学 2023-05-18 Hanna Ziesche , Leonel Rozo

We investigate stochastic combinatorial multi-armed bandit with semi-bandit feedback (CMAB). In CMAB, the question of the existence of an efficient policy with an optimal asymptotic regret (up to a factor poly-logarithmic with the action…

机器学习 · 统计学 2021-01-05 Pierre Perrault , Etienne Boursier , Vianney Perchet , Michal Valko

The restless multi-armed bandit (RMAB) framework is a popular model with applications across a wide variety of fields. However, its solution is hindered by the exponentially growing state space (with respect to the number of arms) and the…

机器学习 · 计算机科学 2025-08-05 Gongpu Chen , Soung Chang Liew , Deniz Gunduz

We establish a unified analytical framework for load balancing systems, which allows us to construct a general class $\Pi$ of policies that are both throughput optimal and heavy-traffic delay optimal. This general class $\Pi$ includes as…

分布式、并行与集群计算 · 计算机科学 2017-10-13 Xingyu Zhou , Fei Wu , Jian Tan , Yin Sun , Ness Shroff