中文
相关论文

相关论文: On Data-Driven Drawdown Control with Restart Mecha…

200 篇论文

In this paper, we consider the design of data-driven predictive controllers for nonlinear systems from input-output data via linear-in-control input Koopman lifted models. Instead of identifying and simulating a Koopman model to predict…

最优化与控制 · 数学 2024-05-03 Thomas de Jong , Valentina Breschi , Maarten Schoukens , Mircea Lazar

Dynamic pricing is both an opportunity and a challenge to the demand side. It is an opportunity as it better reflects the real time market conditions and hence enables an active demand side. However, demand's active participation does not…

系统与控制 · 电气工程与系统科学 2019-12-04 Jiaman Wu , Zhiqi Wang , Chenye Wu , Kui Wang , Yang Yu

The emerging cryptocurrency market has lately received great attention for asset allocation due to its decentralization uniqueness. However, its volatility and brand new trading mode have made it challenging to devising an acceptable…

机器学习 · 计算机科学 2021-10-19 Fengrui Liu , Yang Li , Baitong Li , Jiaxin Li , Huiyang Xie

This study enhances a Deep Q-Network (DQN) trading model by incorporating advanced techniques like Prioritized Experience Replay, Regularized Q-Learning, Noisy Networks, Dueling, and Double DQN. Extensive tests on assets like BTC/USD and…

计算金融 · 定量金融 2023-11-21 Gang Hu

Overconservatism has long been recognized as a major issue with robust optimization, despite its key advantages of tractability, performance guarantee, and limited information. To address this issue, a new criterion is proposed that can…

最优化与控制 · 数学 2026-03-20 Yingjie Lan

Technical trading represents a class of investment strategies for Financial Markets based on the analysis of trends and recurrent patterns of price time series. According standard economical theories these strategies should not be used…

统计金融 · 定量金融 2011-10-25 Federico Garzarelli , Matthieu Cristelli , Andrea Zaccaria , Luciano Pietronero

The need for control strategies that can address dynamic system uncertainty is becoming increasingly important. In this work, we propose a Model Predictive Control by quantifying the risk of failure in our system model. The proposed control…

系统与控制 · 电气工程与系统科学 2023-02-17 Mostafa Tavakkoli Anbarani , Efe C. Balta , Rômulo Meira-Góes , Ilya Kovalenko

Network pruning is an effective method to reduce the computational expense of over-parameterized neural networks for deployment on low-resource systems. Recent state-of-the-art techniques for retraining pruned networks such as weight…

机器学习 · 计算机科学 2021-05-10 Duong H. Le , Binh-Son Hua

In this paper, we propose a novel reinforcement learning algorithm for inventory management of newly launched products with no historical demand information. The algorithm follows the classic Dyna-$Q$ structure, balancing the model-free and…

机器学习 · 计算机科学 2025-06-10 Xinye Qu , Longxiao Liu , Wenjie Huang

This paper studies a data-driven predictive control for a class of control-affine systems which is subject to uncertainty. With the accessibility to finite sample measurements of the uncertain variables, we aim to find controls which are…

最优化与控制 · 数学 2021-05-03 Dan Li , Dariush Fooladivanda , Sonia Martinez

Dynamic decisions are pivotal to economic policy making. We show how existing evidence from randomized control trials can be utilized to guide personalized decisions in challenging dynamic environments with budget and capacity constraints.…

计量经济学 · 经济学 2024-11-26 Karun Adusumilli , Friedrich Geiecke , Claudio Schilter

The democratization of artificial intelligence through decentralized networks represents a paradigm shift in computational provisioning, yet the long-term viability of these ecosystems is critically endangered by the extreme volatility of…

计算机科学与博弈论 · 计算机科学 2026-01-16 Zehua Cheng , Wei Dai , Zhipeng Wang , Rui Sun , Nick Wen , Jiahao Sun

The regression discontinuity (RD) design is widely used for program evaluation with observational data. The primary focus of the existing literature has been the estimation of the local average treatment effect at the existing treatment…

统计方法学 · 统计学 2024-09-05 Yi Zhang , Eli Ben-Michael , Kosuke Imai

Algorithmic trading relies on extracting meaningful signals from diverse financial data sources, including candlestick charts, order statistics on put and canceled orders, traded volume data, limit order books, and news flow. While deep…

机器学习 · 计算机科学 2025-04-22 Kasymkhan Khubiev , Mikhail Semenov

We consider a diffusion risk model where dividends are paid at rate $U(t) \in [0, u_0]$. We are interested in maximising the dividend payments under a drawdown constraint, that is, we penalise a drawdown size larger than a level $d > 0$. We…

最优化与控制 · 数学 2025-11-06 Kira Dudziak , Hanspeter Schmidli

Reinforcement learning is applied to the development of control strategies in order to reduce skin friction drag in a fully developed turbulent channel flow at a low Reynolds number. Motivated by the so-called opposition control (Choi et…

流体动力学 · 物理学 2023-04-26 Takahiro Sonoda , Zhuchen Liu , Toshitaka Itoh , Yosuke Hasegawa

We propose a novel framework for learning stabilizable nonlinear dynamical systems for continuous control tasks in robotics. The key idea is to develop a new control-theoretic regularizer for dynamics fitting rooted in the notion of…

系统与控制 · 计算机科学 2018-11-13 Sumeet Singh , Vikas Sindhwani , Jean-Jacques E. Slotine , Marco Pavone

We consider model-free reinforcement learning (RL) in non-stationary Markov decision processes. Both the reward functions and the state transition functions are allowed to vary arbitrarily over time as long as their cumulative variations do…

机器学习 · 计算机科学 2022-08-23 Weichao Mao , Kaiqing Zhang , Ruihao Zhu , David Simchi-Levi , Tamer Başar

We present a data-driven optimization approach for robotic controlled deposition with a degradable tool. Existing methods make the assumption that the tool tip is not changing or is replaced frequently. Errors can accumulate over time as…

机器人学 · 计算机科学 2023-05-29 Tony Zheng , Monimoy Bujarbaruah , Francesco Borrelli

In this paper, we explore the interplay between Predictive Control and closed-loop optimality, spanning from Model Predictive Control to Data-Driven Predictive Control. Predictive Control in general relies on some form of prediction scheme…

最优化与控制 · 数学 2024-05-29 Akhil S Anand , Shambhuraj Sawant , Dirk Reinhardt , Sebastien Gros