中文
相关论文

相关论文: Deep Reinforcement Learning in Finite-Horizon to E…

200 篇论文

Understanding the transition events between metastable states in complex systems is an important subject in the fields of computational physics, chemistry and biology. The transition pathway plays an important role in characterizing the…

计算物理 · 物理学 2024-04-10 Bo Lin , Yangzheng Zhong , Weiqing Ren

The emergence of transition phenomena between metastable states induced by noise plays a fundamental role in a broad range of nonlinear systems. The computation of the most probable paths is a key issue to understand the mechanism of…

动力系统 · 数学 2021-01-27 Yang Li , Jinqiao Duan , Xianbin Liu

This paper establishes an indirect approximation theorem for the most probable transition pathway of a stochastic interacting particle system in the mean-field framework. This paper studied the problem of indirect approximation of the most…

动力系统 · 数学 2026-05-27 Jianyu Chen , Ting Gao , Galina Strelkova , Jinqiao Duan

In this paper, we investigate the obstacle avoidance and navigation problem in the robotic control area. For solving such a problem, we propose revised Deep Deterministic Policy Gradient (DDPG) and Proximal Policy Optimization algorithms…

机器人学 · 计算机科学 2020-04-13 Daniel Zhang , Colleen P. Bailey

Many natural systems exhibit phase transition where external environmental conditions spark a shift to a new and sometimes quite different state. Therefore, detecting the behavior of a stochastic dynamic system such as the most probable…

最优化与控制 · 数学 2023-03-02 Jianyu Chen , Ting Gao , Yang Li , Jinqiao Duan

Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge. While traditional optimal control methods can find ideal paths, the computational time is often too slow for real-time decision-making. To solve…

最优化与控制 · 数学 2026-04-15 Qiang Le , Yaguang Yang , Isaac E. Weintraub

Due to the sparse rewards and high degree of environment variation, reinforcement learning approaches such as Deep Deterministic Policy Gradient (DDPG) are plagued by issues of high variance when applied in complex real world environments.…

机器人学 · 计算机科学 2018-11-28 Linhai Xie , Yishu Miao , Sen Wang , Phil Blunsom , Zhihua Wang , Changhao Chen , Andrew Markham , Niki Trigoni

This paper tackles the challenge of learning non-Markovian optimal execution strategies in dynamic financial markets. We introduce a novel actor-critic algorithm based on Deep Deterministic Policy Gradient (DDPG) to address this issue, with…

机器学习 · 计算机科学 2024-10-18 Alessandro Micheli , Mélodie Monod

The most probable transition paths of a stochastic dynamical system are the global minimizers of the Onsager-Machlup action functional and can be described by a necessary but not sufficient condition, the Euler-Lagrange equation (a…

数学物理 · 物理学 2023-12-07 Yuanfei Huang , Qiao Huang , Jinqiao Duan

This work is devoted to the investigation of the most probable transition time between metastable states for stochastic dynamical systems. Such a system is modeled by a stochastic differential equation with non-vanishing Brownian noise, and…

数学物理 · 物理学 2021-08-11 Yuanfei Huang , Ying Chao , Wei Wei , Jinqiao Duan

Reinforcement Learning (RL) applied to financial problems has been the subject of a lively area of research. The use of RL for optimal trading strategies that exploit latent information in the market is, to the best of our knowledge, not…

交易与市场微观结构 · 定量金融 2025-11-04 Andrea Macrì , Sebastian Jaimungal , Fabrizio Lillo

Policy optimization is among the most popular and successful reinforcement learning algorithms, and there is increasing interest in understanding its theoretical guarantees. In this work, we initiate the study of policy optimization for the…

机器学习 · 计算机科学 2022-02-08 Liyu Chen , Haipeng Luo , Aviv Rosenberg

Inefficient traffic signal control methods may cause numerous problems, such as traffic congestion and waste of energy. Reinforcement learning (RL) is a trending data-driven approach for adaptive traffic signal control in complex urban…

信号处理 · 电气工程与系统科学 2021-07-14 Zhenning Li , Chengzhong Xu , Guohui Zhang

Lane change is a challenging task which requires delicate actions to ensure safety and comfort. Some recent studies have attempted to solve the lane-change control problem with Reinforcement Learning (RL), yet the action is confined to…

机器人学 · 计算机科学 2019-06-07 Pin Wang , Hanhan Li , Ching-Yao Chan

Extracting governing stochastic differential equation models from elusive data is crucial to understand and forecast dynamics for complex systems. We devise a method to extract the drift term and estimate the diffusion coefficient of a…

数值分析 · 数学 2020-08-21 Jian Ren , Jinqiao Duan

The theory of continuous-time reinforcement learning (RL) has progressed rapidly in recent years. While the ultimate objective of RL is typically to learn deterministic control policies, most existing continuous-time RL methods rely on…

机器学习 · 计算机科学 2026-03-17 Ziheng Cheng , Xin Guo , Yufei Zhang

Deep Reinforcement Learning (DRL) is regarded as a potential method for car-following control and has been mostly studied to support a single following vehicle. However, it is more challenging to learn a stable and efficient car-following…

系统与控制 · 电气工程与系统科学 2022-11-21 Tong Liu , Lei Lei , Kan Zheng , Kuan Zhang

Unmanned aerial vehicles (UAVs) are envisioned to complement the 5G communication infrastructure in future smart cities. Hot spots easily appear in road intersections, where effective communication among vehicles is challenging. UAVs may…

机器学习 · 计算机科学 2023-02-22 Ming Zhu , Xiao-Yang Liu , Anwar Walid

Deep learning and reinforcement learning methods have recently been used to solve a variety of problems in continuous control domains. An obvious application of these techniques is dexterous manipulation tasks in robotics which are…

In this study, we develop a stochastic optimal control approach with reinforcement learning structure to learn the unknown parameters appeared in the drift and diffusion terms of the stochastic differential equation. By choosing an…

最优化与控制 · 数学 2023-08-22 Shuzhen Yang
‹ 上一页 1 2 3 10 下一页 ›