English
Related papers

Related papers: Data-Driven Long-Term Asset Allocation with Tsalli…

200 papers

Trajectory optimization and model predictive control are essential techniques underpinning advanced robotic applications, ranging from autonomous driving to full-body humanoid control. State-of-the-art algorithms have focused on data-driven…

Systems and Control · Electrical Eng. & Systems 2021-11-15 Hany Abdulsamad , Tim Dorau , Boris Belousov , Jia-Jie Zhu , Jan Peters

The data-driven linear quadratic regulator (ddLQR) is a widely studied control method for unknown dynamical systems with disturbance. Existing approaches, both indirect, i.e., those that identify a model followed by model-based design, and…

Optimization and Control · Mathematics 2026-04-13 Thierry Schwaller , Feiran Zhao , Florian Dörfler

We consider a continuous time linear multi inventory system with unknown demands bounded within ellipsoids and controls bounded within ellipsoids or polytopes. We address the problem of "-stabilizing the inventory since this implies some…

Optimization and Control · Mathematics 2007-10-26 D. Bauso , L. Giarré , R. Pesenti

A large class of technically non-chaotic systems, involving scatterings of light particles by flat surfaces with sharp boundaries, is nonetheless characterized by complex random looking motion in phase space. For these systems one may…

Chaotic Dynamics · Physics 2009-11-10 Henk van Beijeren

Distributions derived from non-extensive Tsallis statistics are closely connected with dynamics described by a nonlinear Fokker-Planck equation. The combination shows promise in describing stochastic processes with power-law distributions…

Statistical Mechanics · Physics 2008-12-02 Fredrick Michael , M. D. Johnson

This paper studies a continuous-time stochastic linear-quadratic (SLQ) optimal control problem on infinite-horizon. A data-driven policy iteration algorithm is proposed to solve the SLQ problem. Without knowing three system coefficient…

Optimization and Control · Mathematics 2022-09-30 Heng Zhang , Na Li

This paper studies uniform stabilization and social optimality for linear quadratic (LQ) mean field control problems with multiplicative noise, where agents are coupled via dynamics and individual costs. The state and control weights in…

Optimization and Control · Mathematics 2022-03-31 Bingchang Wang , Huanshui Zhang

Maximum Tsallis entropy (MTE) framework in reinforcement learning has gained popularity recently by virtue of its flexible modeling choices including the widely used Shannon entropy and sparse entropy. However, non-Shannon entropies suffer…

Machine Learning · Computer Science 2022-05-18 Lingwei Zhu , Zheng Chen , Eiji Uchibe , Takamitsu Matsubara

Unlike traditional model-based reinforcement learning approaches that estimate system parameters from data, non-model-based data-driven control learns the optimal policy directly from input-state data without any intermediate model…

Optimization and Control · Mathematics 2026-05-05 Leilei Cui , Zhong-Ping Jiang , Petter N. Kolm , Grégoire G. Macqueron

Temporal distribution shift (TDS) erodes the long-term accuracy of recommender systems, yet industrial practice still relies on periodic incremental training, which struggles to capture both stable and transient patterns. Existing…

Machine Learning · Computer Science 2025-11-27 Yuxuan Zhu , Cong Fu , Yabo Ni , Anxiang Zeng , Yuan Fang

For deterministic continuous time nonlinear control systems, epsilon-practical stabilization entropy and practical stabilization entropy are introduced. Here the rate of attraction is specified by a KL-function. Upper and lower bounds for…

Optimization and Control · Mathematics 2022-12-13 Fritz Colonius , Boumediene Hamzi

Predictive inference requires balancing statistical accuracy against informational complexity, yet the choice of complexity measure is usually imposed rather than derived. We treat econometric objects as predictive rules, mappings from…

Statistics Theory · Mathematics 2026-02-16 Nicholas G. Polson , Daniel Zantedeschi

We propose a data-driven Neural Network (NN) optimization framework to determine the optimal multi-period dynamic asset allocation strategy for outperforming a general stochastic target. We formulate the problem as an optimal stochastic…

Computational Finance · Quantitative Finance 2020-06-30 Chendi Ni , Yuying Li , Peter Forsyth , Ray Carroll

The q-exponential distributions, which are generalizations of the Zipf-Mandelbrot power-law distribution, are frequently encountered in complex systems at their stationary states. From the viewpoint of the principle of maximum entropy, they…

Statistical Mechanics · Physics 2009-11-07 Sumiyoshi Abe

In this paper, we formulate a general time-inconsistent stochastic linear--quadratic (LQ) control problem. The time-inconsistency arises from the presence of a quadratic term of the expected state as well as a state-dependent term in the…

Optimization and Control · Mathematics 2011-11-04 Ying Hu , Hanqing Jin , Xun Yu Zhou

We consider the problem of stochastic optimal control, where the state-feedback control policies take the form of a probability distribution and where a penalty on the entropy is added. By viewing the cost function as a Kullback- Leibler…

Optimization and Control · Mathematics 2024-12-12 Marc Lambert , Francis Bach , Silvère Bonnabel

We apply Tsallis's q-indexed nonextensive entropy to formulate a random matrix theory (RMT), which may be suitable for systems with mixed regular-chaotic dynamics. We consider the super-extensive regime of q < 1. We obtain analytical…

Mathematical Physics · Physics 2011-12-06 A. Abd El-Hady , A. Y. Abul-Magd

We consider an agent trying to bring a system to an acceptable state by repeated probabilistic action. Several recent works on algorithmizations of the Lovasz Local Lemma (LLL) can be seen as establishing sufficient conditions for the agent…

Discrete Mathematics · Computer Science 2016-11-29 Dimitris Achlioptas , Fotis Iliopoulos , Nikos Vlassis

We develop a dynamic trading strategy in the Linear Quadratic Regulator (LQR) framework. By including a price mean-reversion signal into the optimization program, in a trading environment where market impact is linear and stage costs are…

Statistics Theory · Mathematics 2021-11-04 Simon Clinet , Jean-François Perreton , Serge Reydellet

Optimism in the face of uncertainty is a popular approach to balance exploration and exploitation in reinforcement learning. Here, we consider the online linear quadratic regulator (LQR) problem, i.e., to learn the LQR corresponding to an…

Systems and Control · Electrical Eng. & Systems 2026-04-01 Marcell Bartos , Bruce D. Lee , Lenart Treven , Andreas Krause , Florian Dörfler , Melanie N. Zeilinger