中文
相关论文

相关论文: Exploiting random lead times for significant inven…

200 篇论文

This work is motivated by our collaboration with a large consumer packaged goods (CPG) company. We have found that while the company appreciates the advantages of dynamic pricing, they deem it operationally much easier to plan out a static…

数据结构与算法 · 计算机科学 2020-11-24 Will Ma , David Simchi-Levi , Jinglong Zhao

Industrial refrigeration systems have substantial energy needs, but optimizing their operation remains challenging due to the tension between minimizing energy costs and meeting strict cooling requirements. Load shifting--strategic…

最优化与控制 · 数学 2025-10-07 Vade Shah , Yohan John , Ethan Freifeld , Lily Y. Chen , Jason R. Marden

This paper implements the Deep Deterministic Policy Gradient (DDPG) algorithm for computing optimal policies for partially observable single-product periodic review inventory control problems with setup costs and backorders. The decision…

最优化与控制 · 数学 2025-07-29 Eugene Feinberg , Jefferson Huang , Pavlo Kasyanov , Thomas O'Neill

Inventory control is subject to service-level requirements, in which sufficient stock levels must be maintained despite an unknown demand. We propose a data-driven order policy that certifies any prescribed service level under minimal…

机器学习 · 统计学 2024-05-27 Ludvig Hult , Dave Zachariah , Petre Stoica

Standard reinforcement learning (RL) aims to find an optimal policy that identifies the best action for each state. However, in healthcare settings, many actions may be near-equivalent with respect to the reward (e.g., survival). We…

机器学习 · 计算机科学 2020-07-27 Shengpu Tang , Aditya Modi , Michael W. Sjoding , Jenna Wiens

Considering the close interaction between spare parts logistics and maintenance planning, this paper presents a model for joint optimization of multi-location spare parts supply chain and condition-based maintenance under predictive and…

最优化与控制 · 数学 2018-10-17 Morteza Soltani

Adopting a probabilistic approach we determine the optimal dividend payout policy of a firm whose surplus process follows a controlled arithmetic Brownian motion and whose cash-flows are discounted at a stochastic dynamic rate. Dividends…

最优化与控制 · 数学 2021-06-22 Elena Bandini , Tiziano De Angelis , Giorgio Ferrari , Fausto Gozzi

A classical inventory problem is studied from the perspective of embedded options, reducing inventory-management to the design of optimal contracts for forward delivery of stock (commodity). Financial option techniques \`{a} la…

最优化与控制 · 数学 2019-04-10 Roy O. Davies , A. J. Ostaszewski

This paper extends the single-item single-stocking location non-stationary stochastic inventory problem to relax the assumption of independent demand. We present a mathematical programming-based solution method that relaxes the assumption…

最优化与控制 · 数学 2023-09-26 Mengyuan Xiang , Roberto Rossi , Belen Martin-Barragan , S. Armagan Tarim

In this paper, we consider a finite horizon non-stationary inventory system with setup costs. We detail an algorithm to find an approximate optimal policy and the basic idea is to use numerical procedure for computing integrals involved in…

最优化与控制 · 数学 2023-03-16 Jianyong Liu , Wei Geng , Xiaobo Zhao

This paper studies an open question in the warehouse problem where a merchant trading a commodity tries to find an optimal inventory-trading policy to decide on purchase and sale quantities during a fixed time horizon in order to maximize…

数据结构与算法 · 计算机科学 2023-02-24 Ishan Bansal , Oktay Günlük

Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Central to these approaches is the design of the importance sampling (IS) ratio used in…

机器学习 · 计算机科学 2026-05-11 Yuheng Zhang , Chenlu Ye , Shuowei Jin , Changlong Yu , Wei Xiong , Saurabh Sahu , Nan Jiang

We address the problem of finding an optimal policy in a Markov decision process under a restricted policy class defined by the convex hull of a set of base policies. This problem is of great interest in applications in which a number of…

机器学习 · 计算机科学 2018-02-28 Ershad Banijamali , Yasin Abbasi-Yadkori , Mohammad Ghavamzadeh , Nikos Vlassis

We study the problem of forecasting the number of units fulfilled (or ``drained'') from each inventory warehouse to meet customer demand, along with the associated outbound shipping costs. The actual drain and shipping costs are determined…

机器学习 · 计算机科学 2025-07-16 Riccardo Savorgnan , Udaya Ghai , Carson Eisenach , Dean Foster

We study portfolio selection in a complete continuous-time market where the preference is dictated by the rank-dependent utility. As such a model is inherently time inconsistent due to the underlying probability weighting, we study the…

数理金融 · 定量金融 2020-06-04 Ying Hu , Hanqing Jin , Xun Yu Zhou

The beer game is a widely used in-class game that is played in supply chain management classes to demonstrate the bullwhip effect. The game is a decentralized, multi-agent, cooperative problem that can be modeled as a serial supply chain…

机器学习 · 计算机科学 2020-10-15 Afshin Oroojlooyjadid , MohammadReza Nazari , Lawrence Snyder , Martin Takáč

This paper presents a stochastic model predictive control approach for nonlinear systems subject to time-invariant probabilistic uncertainties in model parameters and initial conditions. The stochastic optimal control problem entails a cost…

最优化与控制 · 数学 2014-10-17 Stefan Streif , Matthias Karl , Ali Mesbah

A heterogenous network with base stations (BSs), small base stations (SBSs) and users distributed according to independent Poisson point processes is considered. SBS nodes are assumed to possess high storage capacity and to form a…

信息论 · 计算机科学 2016-11-17 B. N. Bharath , K. G. Nagananda , H. Vincent Poor

In this paper, we develop mixed integer linear programming models to compute near-optimal policy parameters for the non-stationary stochastic lot sizing problem under Bookbinder and Tan's static-dynamic uncertainty strategy. Our models…

最优化与控制 · 数学 2014-09-18 Roberto Rossi , Onur A. Kilic , S. Armagan Tarim

We consider a model where an agent has a repeated decision to make and wishes to maximize their total payoff. Payoffs are influenced by an action taken by the agent, but also an unknown state of the world that evolves over time. Before…

计算机科学与博弈论 · 计算机科学 2021-01-20 Nicole Immorlica , Ian Kash , Brendan Lucier
‹ 上一页 1 8 9 10 下一页 ›