English
Related papers

Related papers: Data-Driven Long-Term Asset Allocation with Tsalli…

200 papers

A novel method of an adaptive linear quadratic (LQ) regulation of uncertain continuous linear time-invariant systems is proposed. Such an approach is based on the direct self-tuning regulators design framework and the exponentially stable…

Systems and Control · Electrical Eng. & Systems 2023-08-22 Anton Glushchenko , Konstantin Lastochkin

Long-term training of large language models (LLMs) requires maintaining stable exploration to prevent the model from collapsing into sub-optimal behaviors. Entropy is crucial in this context, as it controls exploration and helps avoid…

Machine Learning · Computer Science 2026-02-03 Kai Yang , Xin Xu , Yangkun Chen , Weijie Liu , Jiafei Lyu , Zichuan Lin , Deheng Ye , Saiyong Yang

While the idea of robust dynamic programming (DP) is compelling for systems affected by uncertainty, addressing worst-case disturbances generally results in excessive conservatism. This paper introduces a method for constructing control…

Systems and Control · Electrical Eng. & Systems 2025-05-19 Menno van Zutphen , Domagoj Herceg , Duarte J. Antunes

Control of linear dynamics with multiplicative noise naturally introduces robustness against dynamical uncertainty. Moreover, many physical systems are subject to multiplicative disturbances. In this work we show how these dynamics can be…

Optimization and Control · Mathematics 2023-12-27 Peter Coppens , Panagiotis Patrinos

The linear quadratic regulator (LQR) problem has reemerged as an important theoretical benchmark for reinforcement learning-based control of complex dynamical systems with continuous state and action spaces. In contrast with nearly all…

Machine Learning · Computer Science 2020-05-04 Benjamin Gravell , Peyman Mohajerin Esfahani , Tyler Summers

The recently successful Munchausen Reinforcement Learning (M-RL) features implicit Kullback-Leibler (KL) regularization by augmenting the reward function with logarithm of the current stochastic policy. Though significant improvement has…

Machine Learning · Computer Science 2022-05-17 Lingwei Zhu , Zheng Chen , Eiji Uchibe , Takamitsu Matsubara

We consider reinforcement learning (RL) in continuous time and study the problem of achieving the best trade-off between exploration of a black box environment and exploitation of current knowledge. We propose an entropy-regularized reward…

Optimization and Control · Mathematics 2019-02-14 Haoran Wang , Thaleia Zariphopoulou , Xunyu Zhou

The goal of this paper is to develop data-driven control design and evaluation strategies based on linear matrix inequalities (LMIs) and dynamic programming. We consider deterministic discrete-time LTI systems, where the system model is…

Optimization and Control · Mathematics 2021-06-17 Donghwan Lee , Do Wan Kim

We study the discrete-time linear-quadratic (LQ) control model using reinforcement learning (RL). Using entropy to measure the cost of exploration, we prove that the optimal feedback policy for the problem must be Gaussian type. Then, we…

Machine Learning · Statistics 2025-02-05 Lucky Li

This paper presents a novel direct data-driven control framework for solving the linear quadratic regulator (LQR) under disturbances and noisy state measurements. The system dynamics are assumed unknown, and the LQR solution is learned…

Systems and Control · Electrical Eng. & Systems 2025-05-13 Ramin Esmzad , Gokul S. Sankar , Teawon Han , Hamidreza Modares

The principal task to control dynamical systems is to ensure their stability. When the system is unknown, robust approaches are promising since they aim to stabilize a large set of plausible systems simultaneously. We study linear…

Systems and Control · Electrical Eng. & Systems 2020-11-24 Lenart Treven , Sebastian Curi , Mojmir Mutny , Andreas Krause

An amended MaxEnt formulation for systems displaced from the conventional MaxEnt equilibrium is proposed. This formulation involves the minimization of the Kullback-Leibler divergence to a reference $Q$ (or maximization of Shannon…

Mathematical Physics · Physics 2009-11-11 Jean-François Bercher

In this paper we introduce an easy to compute upper bound on the Tsallis entropy of a density matrix describing a system coupled to a noise source. This suggests that the Tsallis entropy is most natural in the context of quantum information…

Quantum Physics · Physics 2017-02-28 Boaz Tamir

Reinforcement learning (RL) has become a key approach for enhancing reasoning in large language models (LLMs), yet scalable training is often hindered by the rapid collapse of policy entropy, which leads to premature convergence and…

Machine Learning · Computer Science 2026-04-14 Ming Lei , Christophe Baehr

We present a sampling-based trajectory optimization method derived from the maximum entropy formulation of Differential Dynamic Programming with Tsallis entropy. This method is a generalization of the legacy work with Shannon entropy, which…

Optimization and Control · Mathematics 2024-09-18 Yuichiro Aoyama , Evangelos A. Theodorou

We study the problem of learning to stabilize (LTS) a linear time-invariant (LTI) system. Policy gradient (PG) methods for control assume access to an initial stabilizing policy. However, designing such a policy for an unknown system is one…

Machine Learning · Computer Science 2025-05-07 Leonardo F. Toso , Lintao Ye , James Anderson

We investigate the cumulative Tsallis entropy, an information measure recently introduced as a cumulative version of the classical Tsallis differential entropy, which is itself a generalization of the Boltzmann-Gibbs statistics. This…

Statistics Theory · Mathematics 2023-06-02 Guillaume Dulac , Thomas Simon

In recent years, the so-called `direct data-driven control' has been a topic of intense research, and it is expected that it will become prominent in future complex dynamical systems control. Within this framework, regularization not only…

Optimization and Control · Mathematics 2026-04-28 Shuyuan Zhang , Zheming Wang , Raphael M. Jungers

Tsallis' non-extensive entropy $S_q$ enables us to treat both a power and exponential evolutions of underlying microscopic dynamics on equal footing by adjusting the variable entropic index $q$ to proper one $q^*$. We propose an alternative…

Statistical Mechanics · Physics 2009-11-07 Wada Tatsuaki , Saito Takeshi

The Tsallis entropy, which is a generalization of the Boltzmann-Gibbs entropy, plays a central role in nonextensive statistical mechanics of complex systems. A lot of efforts have recently been made on establishing a dynamical foundation…

Statistical Mechanics · Physics 2009-11-11 Sumiyoshi Abe , Yutaka Nakada