English
Related papers

Related papers: A Two-fold Randomization Framework for Impulse Con…

200 papers

This paper studies the stabilization for a kind of linear and impulse control systems in finite-dimensional spaces, where impulse instants appear periodically. We present several characterizations on the stabilization; show how to design…

Optimization and Control · Mathematics 2019-07-11 Shulin Qin , Gengsheng Wang , Huaiqiang Yu

Reinforcement learning (RL) is a class of artificial intelligence algorithms being used to design adaptive optimal controllers through online learning. This paper presents a model-free, real-time, data-efficient Q-learning-based algorithm…

Systems and Control · Electrical Eng. & Systems 2023-10-11 Ali Aalipour , Alireza Khani

The ability to direct a Probabilistic Boolean Network (PBN) to a desired state is important to applications such as targeted therapeutics in cancer biology. Reinforcement Learning (RL) has been proposed as a framework that solves a…

Machine Learning · Computer Science 2022-10-26 Sotiris Moschoyiannis , Evangelos Chatzaroulas , Vytenis Sliogeris , Yuhu Wu

We propose a class of numerical schemes for nonlocal HJB variational inequalities (HJBVIs) with monotone drivers. The solution and free boundary of the HJBVI are constructed from a sequence of penalized equations, for which a continuous…

Numerical Analysis · Mathematics 2018-05-17 Christoph Reisinger , Yufei Zhang

The task of inducing, via continuous static state-feedback control, an asymptotically stable heteroclinic orbit in a nonlinear control system is considered in this paper. The main motivation comes from the problem of ensuring convergence to…

Systems and Control · Electrical Eng. & Systems 2023-02-16 Christian Fredrik Sætre , Anton S. Shiriaev

We consider a robust switching control problem. The controller only observes the evolution of the state process, and thus uses feedback (closed-loop) switching strategies, a non standard class of switching controls introduced in this paper.…

Probability · Mathematics 2016-07-04 Erhan Bayraktar , Andrea Cosso , Huyen Pham

We propose a comprehensive framework for policy gradient methods tailored to continuous time reinforcement learning. This is based on the connection between stochastic control problems and randomised problems, enabling applications across…

Optimization and Control · Mathematics 2024-05-01 Robert Denkert , Huyên Pham , Xavier Warin

This paper is concerned with offline reinforcement learning (RL), which learns using pre-collected data without further exploration. Effective offline RL would be able to accommodate distribution shift and limited data coverage. However,…

Machine Learning · Statistics 2024-03-11 Gen Li , Laixi Shi , Yuxin Chen , Yuejie Chi , Yuting Wei

This paper examines reinforcement learning (RL) in infinite-horizon decision processes with almost-sure safety constraints, crucial for applications like autonomous systems, finance, and resource management. We propose a doubly-regularized…

Machine Learning · Computer Science 2025-09-17 Pekka Malo , Lauri Viitasaari , Antti Suominen , Eeva Vilkkumaa , Olli Tahvonen

This paper studies the continuous-time reinforcement learning for stochastic singular control with the application to an infinite-horizon irreversible reinsurance problems. The singular control is equivalently characterized as a pair of…

Optimization and Control · Mathematics 2025-12-03 Zongxia Liang , Xiaodong Luo , Xiang Yu

Optimization problems characterized by both discrete and continuous variables are common across various disciplines, presenting unique challenges due to their complex solution landscapes and the difficulty of navigating mixed-variable…

Optimization and Control · Mathematics 2024-06-03 Haoyan Zhai , Qianli Hu , Jiangning Chen

We propose a reinforcement learning (RL) framework under a broad class of risk objectives, characterized by convex scoring functions. This class covers many common risk measures, such as variance, Expected Shortfall, entropic Value-at-Risk,…

Mathematical Finance · Quantitative Finance 2025-05-16 Shanyu Han , Yang Liu , Xiang Yu

We study risk-sensitive reinforcement learning (RL), a crucial field due to its ability to enhance decision-making in scenarios where it is essential to manage uncertainty and minimize potential adverse outcomes. Particularly, our work…

Machine Learning · Computer Science 2024-07-11 Dake Zhang , Boxiang Lyu , Shuang Qiu , Mladen Kolar , Tong Zhang

In this paper, we study a time-inconsistent stochastic optimal control problem with a recursive cost functional by a multi-person hierarchical differential game approach. An equilibrium strategy of this problem is constructed and a…

Optimization and Control · Mathematics 2016-06-13 Qingmeng Wei , Jiongmin Yong , Zhiyong Yu

In recent times, reinforcement learning has produced baffling results when it comes to performing control tasks with highly non-linear systems. The impressive results always outweigh the potential vulnerabilities or uncertainties associated…

Robotics · Computer Science 2023-11-14 Arshad Javeed

In this paper we study regularity estimates for the solution to an obstacle problem arising in stochastic impulse control theory. We prove using elementary methods the known sharp $C_{loc}^{1,1}$ estimate for the solution. The new proof is…

Analysis of PDEs · Mathematics 2016-12-02 Rohit Jain

The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theory. However, these equations become intractable for high-dimensional or nonlinear systems. This…

Artificial Intelligence · Computer Science 2026-05-04 Preston Rozwood , Edward Mehrez , Ludger Paehler , Wen Sun , Steven L. Brunton

We study policy iteration (PI) for deterministic infinite-horizon discounted optimal control problems, whose value function is characterized by a stationary Hamilton--Jacobi--Bellman (HJB) equation. At the PDE level, PI is fundamentally…

Optimization and Control · Mathematics 2026-04-14 Namkyeong Cho , Yeoneung Kim

This paper presents an inverse optimality method to solve the Hamilton-Jacobi-Bellman equation for a class of nonlinear problems for which the cost is quadratic and the dynamics are affine in the input. The method is inverse optimal because…

Optimization and Control · Mathematics 2011-10-11 Luis Rodrigues , Didier Henrion , Mehdi Abedinpour Fallah

We study a speculative trading problem within the exploratory reinforcement learning (RL) framework of Wang et al. [2020]. The problem is formulated as a sequential optimal stopping problem over entry and exit times under general utility…

Mathematical Finance · Quantitative Finance 2026-04-03 Yun Zhao , Alex S. L. Tse , Harry Zheng
‹ Prev 1 3 4 5 6 7 10 Next ›