中文
相关论文

相关论文: Learning an Inventory Control Policy with General …

200 篇论文

Batch reinforcement learning enables policy learning without direct interaction with the environment during training, relying exclusively on previously collected sets of interactions. This approach is, therefore, well-suited for high-risk…

机器学习 · 计算机科学 2024-11-18 Amna Najib , Stefan Depeweg , Phillip Swazinna

Most existing literature on supply chain and inventory management consider stochastic demand processes with zero or constant lead times. While it is true that in certain niche scenarios, uncertainty in lead times can be ignored, most…

机器学习 · 计算机科学 2022-03-10 Hardik Meisheri , Somjit Nath , Mayank Baranwal , Harshad Khadilkar

We consider a queueing system composed of a dispatcher that routes deterministically jobs to a set of non-observable queues working in parallel. In this setting, the fundamental problem is which policy should the dispatcher implement to…

性能 · 计算机科学 2025-02-23 Jonatha Anselmi , Bruno Gaujal , Tommaso Nesti

We give new approximation algorithms for the submodular joint replenishment problem and the inventory routing problem, using an iterative rounding approach. In both problems, we are given a set of $N$ items and a discrete time horizon of…

数据结构与算法 · 计算机科学 2019-12-03 Thomas Bosman , Neil Olver

This paper studies the automated control method for regulating air conditioner (AC) loads in incentive-based residential demand response (DR). The critical challenge is that the customer responses to load adjustment are uncertain and…

系统与控制 · 电气工程与系统科学 2021-06-15 Xin Chen , Yingying Li , Jun Shimada , Na Li

Learning-based control methods typically assume stationary system dynamics, an assumption often violated in real-world systems due to drift, wear, or changing operating conditions. We study reinforcement learning for control under…

机器学习 · 计算机科学 2026-04-03 Klemens Iten , Bruce Lee , Chenhao Li , Lenart Treven , Andreas Krause , Bhavya Sukhija

We consider the following two deterministic inventory optimization problems over a finite planning horizon $T$ with non-stationary demands. (a) Submodular Joint Replenishment Problem: This involves multiple item types and a single retailer…

数据结构与算法 · 计算机科学 2015-04-27 Viswanath Nagarajan , Cong Shi

Online marketplaces increasingly do more than simply match buyers and sellers: they route orders across competing sellers and, in many categories, offer ancillary fulfillment services that make seller inventory a source of platform revenue.…

多智能体系统 · 计算机科学 2026-04-09 Rene Caldentey , Tong Xie

Overactuated omnidirectional flying vehicles are capable of generating force and torque in any direction, which is important for applications such as contact-based industrial inspection. This comes at the price of an increase in model…

机器人学 · 计算机科学 2020-06-24 Weixuan Zhang , Maximilian Brunner , Lionel Ott , Mina Kamel , Roland Siegwart , Juan Nieto

We consider a distribution logistics scenario where a shipping operator, managing a limited amount of resources, receives a stream of collection requests, issued by a set of customers along a booking time-horizon, that are referred to a…

最优化与控制 · 数学 2023-07-04 Giovanni Giallombardo , Francesca Guerriero , Giovanna Miglionico

In recent years, control methods based on different optimization techniques have shed light on the possibilities of processing information in many quantum systems. When exploring the transmission of quantum states, faster transmission times…

量子物理 · 物理学 2026-01-13 Sofía Perón Santana , Ariel Fiuri , Martín Domínguez , Omar Osenda

This article studies a hyperbolic conservation law that models a highly re-entrant manufacturing system as encountered in semi-conductor production. Characteristic features are the nonlocal character of the velocity and that the influx and…

最优化与控制 · 数学 2009-07-08 Jean-Michel Coron , Matthias kawski , Zhiqiang Wang

The rise of big data analytics has automated the decision-making of companies and increased supply chain agility. In this paper, we study the supply chain contract design problem faced by a data-driven supplier who needs to respond to the…

机器学习 · 计算机科学 2022-11-10 Xuejun Zhao , Ruihao Zhu , William B. Haskell

In many real world problems, control decisions have to be made with limited information. The controller may have no a priori (or even posteriori) data on the nonlinear system, except from a limited number of points that are obtained over…

最优化与控制 · 数学 2011-05-12 Tansu Alpcan

Since its inception in the mid-60s, the inventory staggering problem has been explored and exploited in a wide range of application domains, such as production planning, stock control systems, warehousing, and aerospace/defense logistics.…

数据结构与算法 · 计算机科学 2025-06-13 Noga Alon , Danny Segev

Arrivals in queueing systems are typically assumed to be independent and exponentially distributed. Our analysis of an online bookshop, however, shows that there is an autocorrelation structure present. First, we adjust the inter-arrival…

应用统计 · 统计学 2021-06-29 Petra Tomanová , Vladimír Holý

A key barrier to using reinforcement learning (RL) in many real-world applications is the requirement of a large number of system interactions to learn a good control policy. Off-policy and Offline RL methods have been proposed to reduce…

机器学习 · 计算机科学 2022-12-02 Wenqi Cui , Linbin Huang , Weiwei Yang , Baosen Zhang

Motivated by applications in online marketplaces such as ride-hailing platforms and payment channel networks, we study a single-server queue with state-dependent arrival control. The service operator dynamically chooses the arrival rate as…

最优化与控制 · 数学 2026-02-03 Tianze Qu , Sushil Mahavir Varma

We study a supply chain consisting of production-inventory systems at several locations which are coupled by a common supplier. Demand of customers arrives at each production system according to a Poisson process and is lost if the local…

概率论 · 数学 2023-03-21 Sonja Otten

This letter proposes a novel reinforcement learning method for the synthesis of a control policy satisfying a control specification described by a linear temporal logic formula. We assume that the controlled system is modeled by a Markov…

系统与控制 · 电气工程与系统科学 2020-03-27 Ryohei Oura , Ami Sakakibara , Toshimitsu Ushio