中文
相关论文

相关论文: Learning an Inventory Control Policy with General …

200 篇论文

With the ever increasing prominence of data in retail operations, sales forecasting has become an essential pillar in the efficient management of inventories. When facing high demand, the use of backroom storage and intraday shelf…

应用统计 · 统计学 2019-12-17 Marc-Olivier Boldi , Valérie Chavez-Demoulin , Olivier Gallay

The stochastic knapsack has been used as a model in wide ranging applications from dynamic resource allocation to admission control in telecommunication. In recent years, a variation of the model has become a basic tool in studying problems…

证券定价 · 定量金融 2008-12-02 Grace Lin , Yingdong Lu , David Yao

The Job-Shop Scheduling Problem (JSP) and Flexible Job-Shop Scheduling Problem (FJSP), are canonical combinatorial optimization problems with wide-ranging applications in industrial operations. In recent years, many online reinforcement…

机器学习 · 计算机科学 2025-09-15 Jesse van Remmerden , Zaharah Bukhsh , Yingqian Zhang

Episodic control enables sample efficiency in reinforcement learning by recalling past experiences from an episodic memory. We propose a new model-based episodic memory of trajectories addressing current limitations of episodic control. Our…

机器学习 · 计算机科学 2021-11-09 Hung Le , Thommen Karimpanal George , Majid Abdolshah , Truyen Tran , Svetha Venkatesh

When an item goes out of stock, sales transaction data no longer reflect the original customer demand, since some customers leave with no purchase while others substitute alternative products for the one that was out of stock. Here we…

应用统计 · 统计学 2016-01-15 Benjamin Letham , Lydia M. Letham , Cynthia Rudin

In the context of evolving supply chain management, the significance of efficient inventory management has grown substantially for businesses. However, conventional manual and experience-based approaches often struggle to meet the…

人机交互 · 计算机科学 2025-08-01 Chunan Tong

We present a unified framework for learning continuous control policies using backpropagation. It supports stochastic control by treating stochasticity in the Bellman equation as a deterministic function of exogenous noise. The product is a…

机器学习 · 计算机科学 2015-11-02 Nicolas Heess , Greg Wayne , David Silver , Timothy Lillicrap , Yuval Tassa , Tom Erez

Developing control policies in simulation is often more practical and safer than directly running experiments in the real world. This applies to policies obtained from planning and optimization, and even more so to policies obtained from…

Ensuring quality of service (QoS) guarantees in service systems is a challenging task, particularly when the system is composed of more fine-grained services, such as service function chains. An important QoS metric in service systems is…

性能 · 计算机科学 2020-08-24 Majid Raeis , Ali Tizghadam , Alberto Leon-Garcia

Leveraging machine learning methods to solve constraint satisfaction problems has shown promising, but they are mostly limited to a static situation where the problem description is completely known and fixed from the beginning. In this…

机器学习 · 计算机科学 2025-09-23 Wook Lee , Frans A. Oliehoek

To cope with uncertain traffic patterns and traffic models, traffic-responsive signal control strategies in the literature are designed to be robust to these uncertainties. These robust strategies still require sensing infrastructure to…

系统与控制 · 电气工程与系统科学 2024-10-08 Leonardo Pedroso , Pedro Batista , Markos Papageorgiou

The paper proposes a novel Economic Production Quantity (EPQ) inventory model within a reverse logistics framework, addressing new and repaired products with varying quality and demand patterns. The model integrates production and…

最优化与控制 · 数学 2025-09-25 M. M. Rizvi , I. B. Wadhawan

Inventory models with imperfect quality items are studied by researchers in past two decades. Till now none of them have considered the effect of substitutions to cope up with shortage and avoid lost sales. This paper presents an EOQ…

最优化与控制 · 数学 2018-10-09 Arindum Mukhopadhyay , A. Goswami

We consider the synthesis of control policies from temporal logic specifications for robots that interact with multiple dynamic environment agents. Each environment agent is modeled by a Markov chain whereas the robot is modeled by a finite…

机器人学 · 计算机科学 2012-03-07 Tichakorn Wongpiromsarn , Alphan Ulusoy , Calin Belta , Emilio Frazzoli , Daniela Rus

In many real-world settings, agents must learn from an offline dataset gathered by some prior behavior policy. Such a setting naturally leads to distribution shift between the behavior policy and the target policy being trained - requiring…

We introduce a novel formulation for incorporating visual feedback in controlling robots. We define a generative model from actions to image observations of features on the end-effector. Inference in the model allows us to infer the robot…

机器人学 · 计算机科学 2020-03-11 Nishad Gothoskar , Miguel Lázaro-Gredilla , Abhishek Agarwal , Yasemin Bekiroglu , Dileep George

We study the problem of generating control laws for systems with unknown dynamics. Our approach is to represent the controller and the value function with neural networks, and to train them using loss functions adapted from the…

机器人学 · 计算机科学 2023-02-21 Selim Engin , Volkan Isler

The presented study elaborates a multi-server priority queueing model considering the pre-emptive repeat policy and phase-type distribution (PH) for retrial process. The incoming heterogeneous calls are categorized as handoff calls and new…

最优化与控制 · 数学 2021-08-03 Raina Raj , Vidyottama Jain

This paper demonstrates that continual relearning of control policies using incremental deep reinforcement learning (RL) can improve policy learning for non-stationary processes. We demonstrate this approach for a data-driven 'smart…

机器学习 · 计算机科学 2020-08-06 Avisek Naug , Marcos Quiñones-Grueiro , Gautam Biswas

Managing stock efficiently remains a core issue in modern logistics, where companies must reconcile cost efficiency with dependable service despite unpredictable market conditions. Conventional models often overlook the direct connection…

最优化与控制 · 数学 2026-04-14 Tianxiao Sun , Noah Schwarzkopf
‹ 上一页 1 8 9 10 下一页 ›