中文
相关论文

相关论文: Optimal Stop-Loss and Take-Profit Parameterization…

200 篇论文

Time-distributed Optimization (TDO) is an approach for reducing the computational burden of Model Predictive Control (MPC). When using TDO, optimization iterations are distributed over time by maintaining a running solution estimate and…

最优化与控制 · 数学 2021-02-25 Dominic Liao-McPherson , Terrence Skibik , Jordan Leung , Ilya Kolmanovsky , Marco M. Nicotra

This paper studies spatiotemporal pricing and fleet management for autonomous mobility-on-demand (AMoD) systems while taking elastic demand into account. We consider a platform that offers ride-hailing services using a fleet of autonomous…

最优化与控制 · 数学 2024-04-02 Zhijie Lai , Sen Li

The growth of Robotics-as-a-Service (RaaS) presents new operational challenges, particularly in optimizing business decisions like pricing and equipment management. While much research focuses on the technical aspects of RaaS, the strategic…

最优化与控制 · 数学 2025-10-01 Joo Seung Lee , Anil Aswani

This paper analyzes the timing options embedded in a startup firm, and the associated market entry and exit timing decisions under the exogenous risks of early termination and competitor's entry. Our valuation approach leads to the…

最优化与控制 · 数学 2017-01-13 Tim Leung , Zongxi Li

This study investigates the development of an optimal execution strategy through reinforcement learning, aiming to determine the most effective approach for traders to buy and sell inventory within a finite time horizon. Our proposed model…

交易与市场微观结构 · 定量金融 2025-11-04 Yadh Hafsi , Edoardo Vittori

We consider a stochastic game between a slow institutional investor and a high-frequency trader who are trading a risky asset and their aggregated order-flow impacts the asset price. We model this system by means of two coupled stochastic…

交易与市场微观结构 · 定量金融 2023-06-26 Rama Cont , Alessandro Micheli , Eyal Neuman

This paper studies four trading algorithms of a professional trader at a multilateral trading facility, observing a realistic two-sided limit order book whose dynamics are driven by the order book events. The identity of the trader can be…

交易与市场微观结构 · 定量金融 2015-01-13 Qinghua Li

Online platforms routinely compare multi-armed bandit algorithms, such as UCB and Thompson Sampling, to select the best-performing policy. Unlike standard A/B tests for static treatments, each run of a bandit algorithm over $T$ users…

机器学习 · 计算机科学 2026-04-14 Huiling Meng , Ningyuan Chen , Xuefeng Gao

Fighting Fantasy is a popular recreational fantasy gaming system worldwide. Combat in this system progresses through a stochastic game involving a series of rounds, each of which may be won or lost. Each round, a limited resource (`luck')…

人工智能 · 计算机科学 2020-02-25 Iain G. Johnston

With increasing numbers of mobile robots arriving in real-world applications, more robots coexist in the same space, interact, and possibly collaborate. Methods to provide such systems with system size scalability are known, for example,…

机器人学 · 计算机科学 2024-05-15 Jonas Kuckling , Robin Luckey , Viktor Avrutin , Andrew Vardy , Andreagiovanni Reina , Heiko Hamann

Recent advances in reinforcement learning, such as Dynamic Sampling Policy Optimization (DAPO), show strong performance when paired with large language models (LLMs). Motivated by this success, we ask whether similar gains can be realized…

计算工程、金融与科学 · 计算机科学 2025-05-27 Ruijian Zha , Bojun Liu

We propose and study the integration of sentiment analysis and deep reinforcement learning ensemble algorithms for stock trading by evaluating strategies capable of dynamically altering their active agent given the concurrent market…

交易与市场微观结构 · 定量金融 2024-11-21 Andrew Ye , James Xu , Vidyut Veedgav , Yi Wang , Yifan Yu , Daniel Yan , Ryan Chen , Vipin Chaudhary , Shuai Xu

Many safety-critical real-world problems, such as autonomous driving and collaborative robots, are of a distributed multi-agent nature. To optimize the performance of these systems while ensuring safety, we can cast them as distributed…

系统与控制 · 电气工程与系统科学 2025-08-20 Abdullah Tokmak , Thomas B. Schön , Dominik Baumann

Machine learning and AI-assisted trading have attracted growing interest for the past few years. Here, we use this approach to test the hypothesis that the inefficiency of the cryptocurrency market can be exploited to generate abnormal…

物理与社会 · 物理学 2019-04-09 Laura Alessandretti , Abeer ElBahrawy , Luca Maria Aiello , Andrea Baronchelli

In this paper, we investigate the dynamic emergence of traffic order in a distributed multi-agent system, aiming to minimize inefficiencies that stem from unnecessary structural impositions. We introduce a methodology for developing a…

多智能体系统 · 计算机科学 2025-01-28 Anahita Jain , Husni R. Idris , John-Paul Clarke

One of the main challenges of multi-agent learning lies in establishing convergence of the algorithms, as, in general, a collection of individual, self-serving agents is not guaranteed to converge with their joint policy, when learning…

人工智能 · 计算机科学 2023-05-18 Aleksander Czechowski , Frans A. Oliehoek

We introduce a method to infer lead-lag networks of agents' actions in complex systems. These networks open the way to both microscopic and macroscopic states prediction in such systems. We apply this method to trader-resolved data in the…

交易与市场微观结构 · 定量金融 2018-07-27 Damien Challet , Rémy Chicheportiche , Mehdi Lallouache , Serge Kassibrakis

We present a cross-market algorithmic trading system that balances execution quality with rigorous compliance enforcement. The architecture comprises a high-level planner, a reinforcement learning execution agent, and an independent…

人工智能 · 计算机科学 2025-10-08 Ailiya Borjigin , Cong He

We consider a stochastic lost-sales inventory control system with a lead time $L$ over a planning horizon $T$. Supply is uncertain, and is a function of the order quantity (due to random yield/capacity, etc). We aim to minimize the…

最优化与控制 · 数学 2023-11-01 Boxiao Chen , Jiashuo Jiang , Jiawei Zhang , Zhengyuan Zhou

Reinforcement learning agents for portfolio management are typically trained and deployed as static policies, with no mechanism for using price forecasts at inference time. We propose $\text{FPILOT}$ (**Fin**ancial **P**lugin…

机器学习 · 计算机科学 2026-05-14 Eun Go , Rohan Deb , Arindam Banerjee