中文
相关论文

相关论文: SOLO FTRL algorithm for production management with…

200 篇论文

Resource allocation in distributed and networked systems such as the Cloud is becoming increasingly flexible, allowing these systems to dynamically adjust toward the workloads they serve, in a demand-aware manner. Online balanced…

数据结构与算法 · 计算机科学 2024-10-24 Harald Räcke , Stefan Schmid , Ruslan Zabrodin

Differential machine learning (DML) is a recently proposed technique that uses samplewise state derivatives to regularize least square fits to learn conditional expectations of functionals of stochastic processes as functions of state…

计算金融 · 定量金融 2023-02-21 Arun Kumar Polala , Bernhard Hientzsch

Tabular learning transforms raw features into optimized spaces for downstream tasks, but its effectiveness deteriorates under distribution shifts between training and testing data. We formalize this challenge as the Distribution Shift…

Optimal Order Execution is a well-established problem in finance that pertains to the flawless execution of a trade (buy or sell) for a given volume within a specified time frame. This problem revolves around optimizing returns while…

In collaborative human-robot order picking systems, human pickers and Autonomous Mobile Robots (AMRs) travel independently through a warehouse and meet at pick locations where pickers load items onto the AMRs. In this paper, we consider an…

We propose a model in which dividend payments occur at regular, deterministic intervals in an otherwise continuous model. This contrasts traditional models where either the payment of continuous dividends is controlled or the dynamics are…

最优化与控制 · 数学 2019-07-24 Jussi Keppo , Max Reppen , H. Mete Soner

Anomalous transport in a tilted periodic potential is investigated numerically within the framework of the fractional Fokker-Planck dynamics via the underlying CTRW. An efficient numerical algorithm is developed which is applicable for an…

统计力学 · 物理学 2009-11-11 E. Heinsalu , M. Patriarca , I. Goychuk , G. Schmid , P. Hänggi

Most Reinforcement Learning (RL) methods are traditionally studied in an active learning setting, where agents directly interact with their environments, observe action outcomes, and learn through trial and error. However, allowing…

人工智能 · 计算机科学 2023-10-16 Maryam Zare , Parham M. Kebria , Abbas Khosravi

This paper studies online optimization under inventory (budget) constraints. While online optimization is a well-studied topic, versions with inventory constraints have proven difficult. We consider a formulation of inventory-constrained…

性能 · 计算机科学 2024-12-20 Qiulin Lin , Hanling Yi , John Pang , Minghua Chen , Adam Wierman , Michael Honig , Yuanzhang Xiao

The analysis of markets with indivisible goods and fixed exogenous prices has played an important role in economic models, especially in relation to wage rigidity and unemployment. This research report provides a mathematical and…

综合金融 · 定量金融 2015-08-11 Stefano Nasini , Jordi Castro , Pau Fonseca i Casas

Motivated by practical federated learning settings where clients may not be always available, we investigate a variant of distributed online optimization where agents are active with a known probability $p$ at each time step, and…

机器学习 · 计算机科学 2024-11-26 Juliette Achddou , Nicolò Cesa-Bianchi , Hao Qiu

This paper presents a novel fleet management strategy for battery-powered robot fleets tasked with intra-factory logistics in an autonomous manufacturing facility. In this environment, repetitive material handling operations are subject to…

机器人学 · 计算机科学 2024-09-10 Mithun Goutham , Stephanie Stockar

We study a competitive online optimization problem with multiple inventories. In the problem, an online decision maker seeks to optimize the allocation of multiple capacity-limited inventories over a slotted horizon, while the allocation…

性能 · 计算机科学 2022-02-08 Qiulin Lin , Yanfang Mo , Junyan Su , Minghua Chen

We consider a general class of two-stage distributionally robust optimization (DRO) problems where the ambiguity set is constrained by fixed marginal probability laws that are not necessarily discrete. We derive primal and dual formulations…

最优化与控制 · 数学 2025-10-17 Ariel Neufeld , Qikun Xiang

Reinforcement learning (RL) has shown great effectiveness in quadrotor control, enabling specialized policies to develop even human-champion-level performance in single-task scenarios. However, these specialized policies often struggle with…

机器人学 · 计算机科学 2024-12-18 Jiaxu Xing , Ismail Geles , Yunlong Song , Elie Aljalbout , Davide Scaramuzza

A new approach to obtaining market--directional information, based on a non-stationary solution to the dynamic equation "future price tends to the value that maximizes the number of shares traded per unit time" [1] is presented. In our…

交易与市场微观结构 · 定量金融 2019-05-03 Vladislav Gennadievich Malyshkin

A fractal mobile-immobile (MIM in short) solute transport model in porous media is set forth, and an inverse problem of determining the fractional orders by the additional measurements at one interior point is investigated by Laplace…

数值分析 · 数学 2021-11-29 Gongsheng Li , Xianzheng Jia , Wenyi Liu , Zhiyuan Li

Dynamic pricing strategies are crucial for firms to maximize revenue by adjusting prices based on market conditions and customer characteristics. However, designing optimal pricing strategies becomes challenging when historical data are…

机器学习 · 计算机科学 2025-02-03 Fan Wang , Feiyu Jiang , Zifeng Zhao , Yi Yu

This work studies reinforcement learning (RL) in the context of multi-period supply chains subject to constraints, e.g., on production and inventory. We introduce Distributional Constrained Policy Optimization (DCPO), a novel approach for…

机器学习 · 计算机科学 2023-02-06 Jaime Sabal Bermúdez , Antonio del Rio Chanona , Calvin Tsay

This article develops a deep reinforcement learning (Deep-RL) framework for dynamic pricing on managed lanes with multiple access locations and heterogeneity in travelers' value of time, origin, and destination. This framework relaxes…

系统与控制 · 电气工程与系统科学 2021-01-28 Venktesh Pandey , Evana Wang , Stephen D. Boyles