中文
相关论文

相关论文: Analysis and Evaluation of Baseline Manipulation i…

200 篇论文

In robust Markov decision processes (MDPs), the uncertainty in the transition kernel is addressed by finding a policy that optimizes the worst-case performance over an uncertainty set of MDPs. While much of the literature has focused on…

机器学习 · 计算机科学 2023-03-02 Yue Wang , Alvaro Velasquez , George Atia , Ashley Prater-Bennette , Shaofeng Zou

Resource allocation plays a critical role in minimizing cycle time and improving the efficiency of business processes. Recently, Deep Reinforcement Learning (DRL) has emerged as a powerful technique to optimize resource allocation policies…

机器学习 · 计算机科学 2025-09-03 Jeroen Middelhuis , Zaharah Bukhsh , Ivo Adan , Remco Dijkman

Recommendations Systems allow users to identify trending items among a community while being timely and relevant to the user's expectations. When the purpose of various Recommendation Systems differs, the required type of recommendations…

信息检索 · 计算机科学 2022-05-05 Dinuka Ravijaya Piyadigama , Guhanathan Poravi

We consider data-visualization systems where a middleware layer translates a frontend request to a SQL query to a backend database to compute visual results. We study the problem of answering a visualization request within a limited time…

数据库 · 计算机科学 2022-02-15 Qiushi Bai , Sadeem Alsudais , Chen Li , Shuang Zhao

We study Markov decision processes (MDPs) with multiple limit-average (or mean-payoff) functions. We consider two different objectives, namely, expectation and satisfaction objectives. Given an MDP with k limit-average functions, in the…

计算机科学与博弈论 · 计算机科学 2015-07-01 Tomáš Brázdil , Václav Brožek , Krishnendu Chatterjee , Vojtěch Forejt , Antonín Kučera

Motivated by the operational problems in click and collect systems, such as curbside pickup programs, we study a joint admission control and capacity allocation problem. We consider a system where arriving customers have preferred service…

最优化与控制 · 数学 2022-03-04 Melis Boran , Bahar Cavdar , Tugce Isik

Clinical decision support tools rooted in machine learning and optimization can provide significant value to healthcare providers, including through better management of intensive care units. In particular, it is important that the patient…

机器学习 · 计算机科学 2021-12-20 Fernando Lejarza , Jacob Calvert , Misty M Attwood , Daniel Evans , Qingqing Mao

Offline Reinforcement Learning (RL) aims to learn a near-optimal policy from a fixed dataset of transitions collected by another policy. This problem has attracted a lot of attention recently, but most existing methods with strong…

机器学习 · 计算机科学 2023-05-23 Germano Gabbianelli , Gergely Neu , Nneka Okolo , Matteo Papini

Driven by recent advances in batch Reinforcement Learning (RL), this paper contributes to the application of batch RL to demand response. In contrast to conventional model-based approaches, batch RL techniques do not require a system…

系统与控制 · 计算机科学 2015-04-10 Frederik Ruelens , Bert Claessens , Stijn Vandael , Bart De Schutter , Robert Babuska , Ronnie Belmans

Demand response (DR) is not only a crucial solution to the demand side management but also a vital means of electricity market in maintaining power grid reliability, sustainability and stability. DR can enable consumers (e.g. data centers)…

计算机科学与博弈论 · 计算机科学 2019-01-11 Jianhai Chen , Deshi Ye , Shouling Ji , Qinming He , Yang Xiang , Zhenguang Liu

The classic Dial-A-Ride Problem (DARP) aims at designing the minimum-cost routing that accommodates a set of user requests under constraints at an operations planning level, where users' preferences and revenue management are often…

最优化与控制 · 数学 2020-11-19 Xiaotong Dong , Joseph YJ Chow , S Travis Waller , David Rey

Recently there have been several historical changes in electricity networks that necessitate the development of Demand Side Management (DSM). The main objective of DSM is to achieve an aggregated consumption pattern that is efficient in…

计算机科学与博弈论 · 计算机科学 2019-02-26 Georgios Tsaousoglou , Konstantinos Steriotis , Nikolaos Efthymiopoulos , Konstantinos Smpoukis , Emmanouel Varvarigos

We consider the problem of storing segments of encoded versions of content files in a set of base stations located in a communication cell. These base stations work in conjunction with the main base station of the cell. Users move randomly…

信息论 · 计算机科学 2014-05-27 Konstantinos Poularakis , Leandros Tassiulas

The demand response (DR) program of a traditional HEMS usually intervenes appliances by controlling or scheduling them to achieve multiple objectives such as minimizing energy cost and maximizing user comfort. In this study, instead of…

系统与控制 · 电气工程与系统科学 2021-11-04 Huy Truong Dinh , Kyu-haeng Lee , Daehee Kim

Same-day delivery for e-commerce has become a popular service. Companies usually offer several time delivery options with the earliest one being next hour delivery. Due to tight delivery deadlines and thin margins, companies often find it…

最优化与控制 · 数学 2019-12-09 Anatolii Prokhorchuk , Justin Dauwels , Patrick Jaillet

The planning domain has experienced increased interest in the formal synthesis of decision-making policies. This formal synthesis typically entails finding a policy which satisfies formal specifications in the form of some well-defined…

人工智能 · 计算机科学 2021-11-30 George K. Atia , Andre Beckus , Ismail Alkhouri , Alvaro Velasquez

Nowadays the emerging smart grid technology opens up the possibility of two-way communication between customers and energy utilities. Demand Response Management (DRM) offers the promise of saving money for commercial customers and…

系统与控制 · 电气工程与系统科学 2022-03-07 Hossein Mohammadi Rouzbahani , Abolfazl Rahimnezhad , Hadis Karimipour

In this paper we investigate a dynamic pricing model for constant demand elasticity where customers have a probability distribution on the number of items they order. This is a generalization from standard models which restrict customers to…

最优化与控制 · 数学 2018-03-01 Nyles Breecher , Richard Stockbridge

This paper is concerned with the determination of pricing strategies for a firm that in each period of a finite horizon receives replenishment quantities of a single product which it sells in two markets, e.g., a long-distance market and an…

最优化与控制 · 数学 2015-09-25 Wen , Chen , Adam Fleischhacker , Michael N. Katehakis

A single queue incorporating a retransmission protocol is investigated, assuming that the sequence of per effort success probabilities in the Automatic Retransmission reQuest (ARQ) chain is a priori defined and no channel state information…

多媒体 · 计算机科学 2013-12-03 Anastasios Giovanidis , Gerhard Wunder , Joerg Buehler