English
Related papers

Related papers: An Actor-Critic Framework for Continuous-Time Jump…

200 papers

This paper studies optimal consensus tracking problem of heterogeneous linear multi-agent systems. By introducing tracking error dynamics, the optimal tracking problem is reformulated as finding a Nash-equilibrium solution of a multi-player…

Optimization and Control · Mathematics 2019-05-21 Jilie Zhang , Zhanshan Wang , Hongwei Zhang

We study a time-inconsistent singular stochastic control problem for a general one-dimensional diffusion, where time-inconsistency arises from a non-exponential discount function. To address this, we adopt a game-theoretic framework and…

Optimization and Control · Mathematics 2025-07-08 Andi Bodnariu , Kristoffer Lindensjö , Neofytos Rodosthenous

Reinforcement learning (RL) has proven highly effective in addressing complex decision-making and control tasks. However, in most traditional RL algorithms, the policy is typically parameterized as a diagonal Gaussian distribution with…

Machine Learning · Computer Science 2024-12-24 Yinuo Wang , Likun Wang , Yuxuan Jiang , Wenjun Zou , Tong Liu , Xujie Song , Wenxuan Wang , Liming Xiao , Jiang Wu , Jingliang Duan , Shengbo Eben Li

Soft actor-critic (SAC) is a popular algorithm for max-entropy reinforcement learning. In practice, the energy-based policies in SAC are often approximated using simple policy classes for efficiency, sacrificing the expressiveness and…

Machine Learning · Computer Science 2026-01-01 Yuyang Zhang , Yang Hu , Bo Dai , Na Li

Empirically derived continuum models of collective behavior among large populations of dynamic agents are a subject of intense study in several fields, including biology, engineering and finance. We formulate and study a mean-field game…

Adaptation and Self-Organizing Systems · Physics 2018-06-22 Piyush Grover , Kaivalya Bakshi , Evangelos A. Theodorou

We study time-inconsistent recursive stochastic control problems, i.e., for which the Bellman principle of optimality does not hold. For this class of problems classical optimal controls may fail to exist, or to be relevant in practice, and…

Optimization and Control · Mathematics 2024-03-14 Elisa Mastrogiacomo , Marco Tarsia

In many multi-agent systems, agents interact repeatedly and are expected to settle into stable, rational behavior over time. Yet in practice, behavior often drifts, and detecting such deviations in real time remains an open challenge. We…

Computer Science and Game Theory · Computer Science 2026-05-25 Etienne Gauthier , Francis Bach , Michael I. Jordan

A multi-agent system operates in an uncertain environment about which agents have different and time varying beliefs that, as time progresses, converge to a common belief. A global utility function that depends on the realized state of the…

Computer Science and Game Theory · Computer Science 2016-02-08 Ceyhun Eksin , Alejandro Ribeiro

In this paper, we consider a linear-quadratic optimal control problem of mean-field stochastic differential equation with jump diffusion, which is also called as an MF-LQJ problem. Here, cost functional is allowed to be indefinite. We use…

Optimization and Control · Mathematics 2021-11-18 Guangchen Wang , Wencan Wang

This paper studies a continuous-time market {under stochastic environment} where an agent, having specified an investment horizon and a target terminal mean return, seeks to minimize the variance of the return with multiple stocks and a…

Portfolio Management · Quantitative Finance 2013-02-28 Wan-Kai Pang , Yuan-Hua Ni , Xun Li , Ka-Fai Cedric Yiu

This paper analyzes a class of impulse control problems for multi-dimensional jump diffusions in the finite time horizon. Following the basic mathematical setup from Stroock and Varadhan \cite{StroockVaradhan06}, this paper first…

Optimization and Control · Mathematics 2013-04-23 Yann-Shin Aaron Chen , Xin Guo

This paper studies the stochastic optimal control of jump-diffusion processes and the associated fully nonlinear backward stochastic Hamilton--Jacobi--Bellman (BSHJB) equations. We establish the dynamic programming principle (DPP) via…

Optimization and Control · Mathematics 2026-05-21 Dunxiang Liang , Qingxin Meng

Reaction-diffusion systems driven far from thermodynamic equilibrium through the injection of energy can support multiple distinct spatial patterns that persist as long-lived dynamical phases. The stability of these metastable phases is not…

Statistical Mechanics · Physics 2026-03-13 Eric R. Heller , David T. Limmer

We present a dynamic programming-based solution to a stochastic optimal control problem up to a hitting time for a discrete-time Markov control process. Firstly, we determine an optimal control policy to steer the process toward a compact…

Optimization and Control · Mathematics 2009-09-28 Debasish Chatterjee , Eugenio Cinquemani , Giorgos Chaloulos , John Lygeros

Diffusion- and flow-based policies deliver state-of-the-art performance on long-horizon robotic manipulation and imitation learning tasks. However, these controllers employ a fixed inference budget at every control step, regardless of task…

Robotics · Computer Science 2025-11-27 Inkook Chun , Seungjae Lee , Michael S. Albergo , Saining Xie , Eric Vanden-Eijnden

This paper considers linear-quadratic (LQ) stochastic leader-follower Stackelberg differential games for jump-diffusion systems with random coefficients. We first solve the LQ problem of the follower using the stochastic maximum principle…

Optimization and Control · Mathematics 2020-10-07 Jun Moon

Recent studies have increasingly focused on non-asymptotic convergence analyses for actor-critic (AC) algorithms. One such effort introduced a two-timescale critic-actor algorithm for the discounted cost setting using a tabular…

Machine Learning · Computer Science 2025-10-07 Prashansa Panda , Shalabh Bhatnagar

Motivated by applications in risk-sensitive reinforcement learning, we study mean-variance optimization in a discounted reward Markov Decision Process (MDP). Specifically, we analyze a Temporal Difference (TD) learning algorithm with linear…

Machine Learning · Computer Science 2025-03-13 Tejaram Sangadi , L. A. Prashanth , Krishna Jagannathan

In this paper, a novel decentralized intelligent adaptive optimal strategy has been developed to solve the pursuit-evasion game for massive Multi-Agent Systems (MAS) under uncertain environment. Existing strategies for pursuit-evasion games…

Systems and Control · Electrical Eng. & Systems 2020-08-10 Zejian Zhou , Hao Xu

While deep learning methods have achieved strong performance in time series prediction, their black-box nature and inability to explicitly model underlying stochastic processes often limit their generalization to non-stationary data,…

Machine Learning · Computer Science 2026-02-10 Yuanpei Gao , Qi Yan , Yan Leng , Renjie Liao