English
Related papers

Related papers: Learning an Inventory Control Policy with General …

200 papers

Batch reinforcement learning enables policy learning without direct interaction with the environment during training, relying exclusively on previously collected sets of interactions. This approach is, therefore, well-suited for high-risk…

Machine Learning · Computer Science 2024-11-18 Amna Najib , Stefan Depeweg , Phillip Swazinna

Most existing literature on supply chain and inventory management consider stochastic demand processes with zero or constant lead times. While it is true that in certain niche scenarios, uncertainty in lead times can be ignored, most…

Machine Learning · Computer Science 2022-03-10 Hardik Meisheri , Somjit Nath , Mayank Baranwal , Harshad Khadilkar

We consider a queueing system composed of a dispatcher that routes deterministically jobs to a set of non-observable queues working in parallel. In this setting, the fundamental problem is which policy should the dispatcher implement to…

Performance · Computer Science 2025-02-23 Jonatha Anselmi , Bruno Gaujal , Tommaso Nesti

We give new approximation algorithms for the submodular joint replenishment problem and the inventory routing problem, using an iterative rounding approach. In both problems, we are given a set of $N$ items and a discrete time horizon of…

Data Structures and Algorithms · Computer Science 2019-12-03 Thomas Bosman , Neil Olver

This paper studies the automated control method for regulating air conditioner (AC) loads in incentive-based residential demand response (DR). The critical challenge is that the customer responses to load adjustment are uncertain and…

Systems and Control · Electrical Eng. & Systems 2021-06-15 Xin Chen , Yingying Li , Jun Shimada , Na Li

Learning-based control methods typically assume stationary system dynamics, an assumption often violated in real-world systems due to drift, wear, or changing operating conditions. We study reinforcement learning for control under…

Machine Learning · Computer Science 2026-04-03 Klemens Iten , Bruce Lee , Chenhao Li , Lenart Treven , Andreas Krause , Bhavya Sukhija

We consider the following two deterministic inventory optimization problems over a finite planning horizon $T$ with non-stationary demands. (a) Submodular Joint Replenishment Problem: This involves multiple item types and a single retailer…

Data Structures and Algorithms · Computer Science 2015-04-27 Viswanath Nagarajan , Cong Shi

Online marketplaces increasingly do more than simply match buyers and sellers: they route orders across competing sellers and, in many categories, offer ancillary fulfillment services that make seller inventory a source of platform revenue.…

Multiagent Systems · Computer Science 2026-04-09 Rene Caldentey , Tong Xie

Overactuated omnidirectional flying vehicles are capable of generating force and torque in any direction, which is important for applications such as contact-based industrial inspection. This comes at the price of an increase in model…

Robotics · Computer Science 2020-06-24 Weixuan Zhang , Maximilian Brunner , Lionel Ott , Mina Kamel , Roland Siegwart , Juan Nieto

We consider a distribution logistics scenario where a shipping operator, managing a limited amount of resources, receives a stream of collection requests, issued by a set of customers along a booking time-horizon, that are referred to a…

Optimization and Control · Mathematics 2023-07-04 Giovanni Giallombardo , Francesca Guerriero , Giovanna Miglionico

In recent years, control methods based on different optimization techniques have shed light on the possibilities of processing information in many quantum systems. When exploring the transmission of quantum states, faster transmission times…

Quantum Physics · Physics 2026-01-13 Sofía Perón Santana , Ariel Fiuri , Martín Domínguez , Omar Osenda

This article studies a hyperbolic conservation law that models a highly re-entrant manufacturing system as encountered in semi-conductor production. Characteristic features are the nonlocal character of the velocity and that the influx and…

Optimization and Control · Mathematics 2009-07-08 Jean-Michel Coron , Matthias kawski , Zhiqiang Wang

The rise of big data analytics has automated the decision-making of companies and increased supply chain agility. In this paper, we study the supply chain contract design problem faced by a data-driven supplier who needs to respond to the…

Machine Learning · Computer Science 2022-11-10 Xuejun Zhao , Ruihao Zhu , William B. Haskell

In many real world problems, control decisions have to be made with limited information. The controller may have no a priori (or even posteriori) data on the nonlinear system, except from a limited number of points that are obtained over…

Optimization and Control · Mathematics 2011-05-12 Tansu Alpcan

Since its inception in the mid-60s, the inventory staggering problem has been explored and exploited in a wide range of application domains, such as production planning, stock control systems, warehousing, and aerospace/defense logistics.…

Data Structures and Algorithms · Computer Science 2025-06-13 Noga Alon , Danny Segev

Arrivals in queueing systems are typically assumed to be independent and exponentially distributed. Our analysis of an online bookshop, however, shows that there is an autocorrelation structure present. First, we adjust the inter-arrival…

Applications · Statistics 2021-06-29 Petra Tomanová , Vladimír Holý

A key barrier to using reinforcement learning (RL) in many real-world applications is the requirement of a large number of system interactions to learn a good control policy. Off-policy and Offline RL methods have been proposed to reduce…

Machine Learning · Computer Science 2022-12-02 Wenqi Cui , Linbin Huang , Weiwei Yang , Baosen Zhang

Motivated by applications in online marketplaces such as ride-hailing platforms and payment channel networks, we study a single-server queue with state-dependent arrival control. The service operator dynamically chooses the arrival rate as…

Optimization and Control · Mathematics 2026-02-03 Tianze Qu , Sushil Mahavir Varma

We study a supply chain consisting of production-inventory systems at several locations which are coupled by a common supplier. Demand of customers arrives at each production system according to a Poisson process and is lost if the local…

Probability · Mathematics 2023-03-21 Sonja Otten

This letter proposes a novel reinforcement learning method for the synthesis of a control policy satisfying a control specification described by a linear temporal logic formula. We assume that the controlled system is modeled by a Markov…

Systems and Control · Electrical Eng. & Systems 2020-03-27 Ryohei Oura , Ami Sakakibara , Toshimitsu Ushio
‹ Prev 1 4 5 6 7 8 10 Next ›