中文
相关论文

相关论文: Learning an Inventory Control Policy with General …

200 篇论文

There has been growing interest in applying reinforcement learning (RL) to inventory management, either by optimizing over temporal transitions or by learning directly from full historical demand trajectories. This contrasts sharply with…

机器学习 · 统计学 2026-02-03 Yaqi Xie , Will Ma , Linwei Xin

Standard approaches to sequential decision-making exploit an agent's ability to continually interact with its environment and improve its control policy. However, due to safety, ethical, and practicality constraints, this type of…

机器学习 · 计算机科学 2023-05-11 Patrick Emedom-Nnamdi , Abram L. Friesen , Bobak Shahriari , Nando de Freitas , Matt W. Hoffman

Through the analysis of a dataset of ultra high frequency order book updates, we introduce a model which accommodates the empirical properties of the full order book together with the stylized facts of lower frequency financial data. To do…

交易与市场微观结构 · 定量金融 2014-09-05 Weibing Huang , Charles-Albert Lehalle , Mathieu Rosenbaum

In this paper, we consider the periodic reference tracking problem in the framework of batch-mode reinforcement learning, which studies methods for solving optimal control problems from the sole knowledge of a set of trajectories. In…

系统与控制 · 计算机科学 2013-03-13 Aivar Sootla , Natalja Strelkowa , Damien Ernst , Mauricio Barahona , Guy-Bart Stan

We consider the setting of iterative learning control, or model-based policy learning in the presence of uncertain, time-varying dynamics. In this setting, we propose a new performance metric, planning regret, which replaces the standard…

机器学习 · 计算机科学 2021-03-01 Naman Agarwal , Elad Hazan , Anirudha Majumdar , Karan Singh

We consider a periodic review inventory control problem having an underlying modulation process that affects demand and that is partially observed by the uncensored demand process and a novel additional observation data (AOD) process. We…

最优化与控制 · 数学 2022-02-15 Satya S. Malladi , Alan L. Erera , Chelsea C. White

We introduce a method by which a generative model learning the joint distribution between actions and future states can be used to automatically infer a control scheme for any desired reward function, which may be altered on the fly without…

机器学习 · 计算机科学 2017-03-10 Nicholas Guttenberg , Yen Yu , Ryota Kanai

We consider a stochastic inventory control problem under censored demands, lost sales, and positive lead times. This is a fundamental problem in inventory management, with significant literature establishing near-optimality of a simple…

机器学习 · 计算机科学 2019-05-14 Shipra Agrawal , Randy Jia

In this paper, we study the offline sequential feature-based pricing and inventory control problem where the current demand depends on the past demand levels and any demand exceeding the available inventory is lost. Our goal is to leverage…

机器学习 · 统计学 2026-03-12 Korel Gundem , Zhengling Qi

In this paper, we consider an infinite horizon, continuous-review, stochastic inventory system in which cumulative customers' demand is price-dependent and is modeled as a Brownian motion. Excess demand is backlogged. The revenue is earned…

最优化与控制 · 数学 2018-07-12 Dacheng Yao

We present a control strategy that applies inverse dynamics to a learned acceleration error model for accurate multirotor control input generation. This allows us to retain accurate trajectory and control input generation despite the…

机器人学 · 计算机科学 2020-11-03 Alexander Spitzer , Nathan Michael

Inventory management is a fundamental challenge in supply chain management. The challenge is compounded when the associated products have unpredictable demands. This study proposes an innovative optimization approach combining…

最优化与控制 · 数学 2024-02-20 Sarit Maitra

We study generalizable policy learning from demonstrations for complex low-level control (e.g., contact-rich object manipulations). We propose a novel hierarchical imitation learning method that utilizes sub-optimal demos. Firstly, we…

机器学习 · 计算机科学 2024-07-09 Zhiwei Jia , Vineet Thumuluri , Fangchen Liu , Linghao Chen , Zhiao Huang , Hao Su

Inventory and queueing systems are often designed by controlling weighted combination of some time-averaged performance metrics (like cumulative holding, shortage, server-utilization or congestion costs); but real-world constraints, like…

最优化与控制 · 数学 2025-07-01 Madhu Dhiman , Veeraruna Kavitha , Nandyala Hemachandra

In this paper, we propose an inventory model where items are inspected through multiple screening processes before delivery to customers. Each screening process has independent screening rate and defective percentage. Defective items…

最优化与控制 · 数学 2013-02-07 Allen H. Tai

Closed-loop management of geological CO2 storage requires control policies that adapt to uncertain reservoir behavior while relying on observations that are realistically available during operation. This work formulates CO2 injection and…

机器学习 · 计算机科学 2026-05-05 Sofianos Panagiotis Fotias , Vassilis Gaganis

We consider an inventory system in which inventory level fluctuates as a Brownian motion in the absence of control. The inventory continuously accumulates cost at a rate that is a general convex function of the inventory level, which can be…

最优化与控制 · 数学 2014-01-21 Jim Dai , Dacheng Yao

We consider a discrete-time bipartite matching model with random arrivals of units of supply and demand that can wait in queues located at the nodes in the network. A control policy determines which are matched at each time. The focus is on…

离散数学 · 计算机科学 2016-06-28 Ana Bušić , Sean Meyn

A self-learning optimal control algorithm for episodic fixed-horizon manufacturing processes with time-discrete control actions is proposed and evaluated on a simulated deep drawing process. The control model is built during consecutive…

系统与控制 · 计算机科学 2020-01-07 Johannes Dornheim , Norbert Link , Peter Gumbsch

We consider a continuous-review inventory system in which the setup cost of each order is a general function of the order quantity and the demand process is modeled as a Brownian motion with a positive drift. Assuming the holding and…

最优化与控制 · 数学 2020-09-03 Shuangchi He , Dacheng Yao , Hanqin Zhang