English
Related papers

Related papers: Residual U-net with Self-Attention to Solve Multi-…

200 papers

This paper studies an optimal dividend problem for a company that aims to maximize the mean-variance (MV) objective of the accumulated discounted dividend payments up to its ruin time. The MV objective involves an integral form over a…

Optimization and Control · Mathematics 2025-08-19 Jingyi Cao , Dongchen Li , Virginia R. Young , Bin Zou

Stochastic optimal principle leads to the resolution of a partial differential equation (PDE), namely the Hamilton-Jacobi-Bellman (HJB) equation. In general, this equation cannot be solved analytically, thus numerical algorithms are the…

Numerical Analysis · Mathematics 2021-09-14 Christelle Dleuna Nyoumbi , Antoine Tambue

Multi-agent navigation in unknown and cluttered environments has broad applications, yet remains fundamentally challenging. In particular, dense agent-agent and agent-obstacle reactive interactions can exacerbate the inherent competition…

Systems and Control · Electrical Eng. & Systems 2026-05-14 Fenglan Wang , Xinguo Shu , Lei He , Lin Zhao

In this paper, we introduce Hamilton-Jacobi-Bellman (HJB) equations for Q-functions in continuous time optimal control problems with Lipschitz continuous controls. The standard Q-function used in reinforcement learning is shown to be the…

Optimization and Control · Mathematics 2020-05-05 Jeongho Kim , Insoon Yang

We consider the problem of portfolio optimization in a simple incomplete market and under a general utility function. By working with the associated Hamilton-Jacobi-Bellman partial differential equation (HJB PDE), we obtain a closed-form…

Probability · Mathematics 2018-02-22 Rohini Kumar , Hussein Nasralah

We address the problem of computing a control for a time-dependent nonlinear system to reach a target set in a minimal time. To solve this minimal time control problem, we introduce a hierarchy of linear semi-infinite programs, the values…

Optimization and Control · Mathematics 2023-07-04 Antoine Oustry , Matteo Tacchi

We propose a method to efficiently estimate the eigenvalues of any arbitrary (potentially weighted and/or directed) network of interacting dynamical agents from dynamical observations. These observations are discrete, temporal measurements…

Optimization and Control · Mathematics 2022-03-28 Mikhail Hayhoe , Francisco Barreras , Victor M. Preciado

Recently, a scalable approach to system analysis and controller synthesis for homogeneous multi-agent systems with Bernoulli distributed packet loss has been proposed. As a key result of that line of work, it was shown how to obtain upper…

Systems and Control · Electrical Eng. & Systems 2023-11-28 Christian Hespe , Herbert Werner

Cooperative problems under continuous control have always been the focus of multi-agent reinforcement learning. Existing algorithms suffer from the problem of uneven learning degree with the increase of the number of agents. In this paper,…

Multiagent Systems · Computer Science 2021-07-05 Kai Liu , Yuyang Zhao , Gang Wang , Bei Peng

This paper investigates the optimal control problems for the finite-horizon continuous-time Markov decision processes with delay-dependent control policies. We develop compactification methods in decision processes, and show that the…

Probability · Mathematics 2023-07-06 Zhong-Wei Liao , Jinghai Shao

We study the problem of serving randomly arriving and delay-sensitive traffic over a multi-channel communication system with time-varying channel states and unknown statistics. This problem deviates from the classical…

Networking and Internet Architecture · Computer Science 2018-11-28 Semih Cayci , Atilla Eryilmaz

The main challenge of multiagent reinforcement learning is the difficulty of learning useful policies in the presence of other simultaneously learning agents whose changing behaviors jointly affect the environment's transition and reward…

This paper considers a robust time-consistent mean-variance-skewness portfolio selection problem for an ambiguity-averse investor by taking into account wealth-dependent risk aversion and wealth-dependent skewness preference as well as…

Optimization and Control · Mathematics 2022-01-19 Jian-hao Kang , Nan-jing Huang , Zhihao Hu , Ben-Zhang Yang

We study decentralized multi-agent multi-armed bandits in fully heavy-tailed settings, where clients communicate over sparse random graphs with heavy-tailed degree distributions and observe heavy-tailed (homogeneous or heterogeneous) reward…

Machine Learning · Computer Science 2025-02-03 Xingyu Wang , Mengfan Xu

We address the crucial yet underexplored stability properties of the Hamilton--Jacobi--Bellman (HJB) equation in model-free reinforcement learning contexts, specifically for Lipschitz continuous optimal control problems. We bridge the gap…

Optimization and Control · Mathematics 2024-04-23 Namkyeong Cho , Yeoneung Kim

We study high-dimensional stochastic optimal control problems in which many agents cooperate to minimize a convex cost functional. We consider both the full-information problem, in which each agent observes the states of all other agents,…

Probability · Mathematics 2023-01-10 Joe Jackson , Daniel Lacker

We establish a well-posedness and error-estimation framework that solves Hamilton-Jacobi equations by minimizing the least-squares residual of monotone finite-difference discretizations. This approach also applies naturally to second-order…

Numerical Analysis · Mathematics 2026-05-13 Olivier Bokanowski , Carlos Esteve-Yagüe , Richard Tsai

We study optimal liquidation of a trading position (so-called block order or meta-order) in a market with a linear temporary price impact (Kyle, 1985). We endogenize the pressure to liquidate by introducing a downward drift in the…

Portfolio Management · Quantitative Finance 2018-05-25 Pavol Brunovský , Aleš Černý , Ján Komadel

This paper characterizes differentiable subgame perfect equilibria in a continuous time intertemporal decision optimization problem with non-constant discounting. The equilibrium equation takes two different forms, one of which is…

Optimization and Control · Mathematics 2007-05-23 Ivar Ekeland , Ali Lazrak

The ad hoc coordination problem is to design an autonomous agent which is able to achieve optimal flexibility and efficiency in a multiagent system with no mechanisms for prior coordination. We conceptualise this problem formally using a…

Computer Science and Game Theory · Computer Science 2015-06-04 Stefano V. Albrecht , Subramanian Ramamoorthy
‹ Prev 1 3 4 5 6 7 10 Next ›