English
Related papers

Related papers: Robust Performance Analysis of Source-Seeking Dyna…

200 papers

The nonlinear response coefficient, $\chi_{4,22}$, is a crucial observable for probing the dynamical properties of the quark-gluon plasma (QGP). While traditionally understood as a signature of medium response, recent studies suggest that…

Nuclear Theory · Physics 2026-04-23 Zhi-Jie Yang , Hao-jie Xu , Jie Zhao , Hanlin Li

This paper introduces a novel data-driven approach to design a linear quadratic regulator (LQR) using a reinforcement learning (RL) algorithm that does not require a system model. The key contribution is to perform policy iteration (PI) by…

Systems and Control · Electrical Eng. & Systems 2023-11-20 Soroush Asri , Luis Rodrigues

We study the global convergence of generative adversarial imitation learning for linear quadratic regulators, which is posed as minimax optimization. To address the challenges arising from non-convex-concave geometry, we analyze the…

Machine Learning · Computer Science 2019-01-15 Qi Cai , Mingyi Hong , Yongxin Chen , Zhaoran Wang

The Quark--Meson--Coupling (QMC) model self-consistently relates the dynamics of the internal quark structure of a hadron to the relativistic mean fields arising in nuclear matter. It offers a natural explanation to some open questions in…

The sample inefficiency of reinforcement learning (RL) remains a significant challenge in robotics. RL requires large-scale simulation and can still cause long training times, slowing research and innovation. This issue is particularly…

Robotics · Computer Science 2026-01-16 Johannes Heeg , Yunlong Song , Davide Scaramuzza

Under dynamic traffic, service function chain (SFC) migration is considered as an effective way to improve resource utilization. However, the lack of future network information leads to non-optimal solutions, which motivates us to study…

Networking and Internet Architecture · Computer Science 2019-11-14 Ruoyun Chen , Hancheng Lu , Yujiao Lu , Jinxue Liu

Learning in multi-agent systems is highly challenging due to several factors including the non-stationarity introduced by agents' interactions and the combinatorial nature of their state and action spaces. In particular, we consider the…

Machine Learning · Statistics 2023-05-10 Barna Pásztor , Ilija Bogunovic , Andreas Krause

A scalable and resource-efficient quantum reinforcement learning framework is presented that eliminates the linear qubit-scaling barrier in multi-step quantum Markov decision processes (QMDPs). The proposed framework integrates a QMDP…

Quantum Physics · Physics 2026-04-23 Thet Htar Su , Shaswot Shresthamali , Masaaki Kondo

We introduce an extended nonlinear Lugiato-Lefever equation (LLE) with the pseudo-stimulated-Raman-scattering (pseudo-SRS) cubic term, linear damping/gain, and spatial inhomogeneous (weakly or strongly localized) pump. The LLE is derived,…

Pattern Formation and Solitons · Physics 2025-12-02 Evgeny M. Gromov , Boris A. Malomed

Traffic assignment methods are some of the key approaches used to model flow patterns that arise in transportation networks. Since static traffic assignment does not have a notion of time, it is not designed to represent temporal dynamics…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-05-20 Cy Chan , Anu Kuncheria , Bingyu Zhao , Theophile Cabannes , Alexander Keimer , Bin Wang , Alexandre Bayen , Jane Macfarlane

Reinforcement learning (RL) has shown great effectiveness in quadrotor control, enabling specialized policies to develop even human-champion-level performance in single-task scenarios. However, these specialized policies often struggle with…

Robotics · Computer Science 2024-12-18 Jiaxu Xing , Ismail Geles , Yunlong Song , Elie Aljalbout , Davide Scaramuzza

We propose a method for designing policies for convex stochastic control problems characterized by random linear dynamics and convex stage cost. We consider policies that employ quadratic approximate value functions as a substitute for the…

Optimization and Control · Mathematics 2023-11-10 Alan Yang , Stephen Boyd

Deep reinforcement learning has shown promise in various engineering applications, including vehicular traffic control. The non-stationary nature of traffic, especially in the lane-free environment with more degrees of freedom in vehicle…

Robotics · Computer Science 2024-06-24 Mehran Berahman , Majid Rostami-Shahrbabaki , Klaus Bogenberger

A data-driven framework is proposed for online estimation of quadrotor motor efficiency via residual minimization. The problem is formulated as a constrained nonlinear optimization that minimizes trajectory residuals between measured flight…

Systems and Control · Electrical Eng. & Systems 2026-03-09 Sheng-Wen Cheng , Teng-Hu Cheng

Traffic congestion, primarily driven by intersection queuing, significantly impacts urban living standards, safety, environmental quality, and economic efficiency. While Traffic Signal Control (TSC) systems hold potential for congestion…

Machine Learning · Computer Science 2026-01-14 Qiang Li , Jin Niu , Lina Yu

This paper presents a formulation of Lagrangian dynamics of constrained mechanical systems in terms of reduced quasi-velocities and quasi-forces that can be used for simulation, analysis, and control purposes. In this formulation, Cholesky…

Computational Physics · Physics 2021-08-17 Farhad Aghili

An iterative optimization approach that simultaneously minimizes the energy and optimizes the Lagrange multipliers enforcing desired constraints is presented. The method is tested on previously established benchmark systems and it is proved…

Computational Physics · Physics 2018-08-15 D. Kidd , A. S. Umar , K. Varga

In this paper, we present a computationally efficient trajectory optimizer that can exploit GPUs to jointly compute trajectories of tens of agents in under a second. At the heart of our optimizer is a novel reformulation of the non-convex…

Robotics · Computer Science 2020-11-10 Fatemeh Rastgar , Houman Masnavi , Jatan Shrestha , Karl Kruusamae , Alvo Aabloo , Arun Kumar Singh

We study the long-time dynamics of a bulk-surface convective Cahn--Hilliard system describing phase separation processes with bulk-surface interaction. The presence of convection terms leads to a non-autonomous dynamical system and prevents…

Analysis of PDEs · Mathematics 2026-03-12 Patrik Knopf , Andrea Poiatti , Jonas Stange , Sema Yayla

This article develops a strengthened convex quadratic convex (QC) relaxation of the AC Optimal Power Flow (AC-OPF) problem and presents an optimization-based bound-tightening (OBBT) algorithm to compute tight, feasible bounds on the voltage…

Optimization and Control · Mathematics 2019-01-30 Kaarthik Sundar , Harsha Nagarajan , Sidhant Misra , Mowen Lu , Carleton Coffrin , Russell Bent