English
Related papers

Related papers: Tsallis Entropy Regularization for Linearly Solvab…

200 papers

Despite the many recent advances in reinforcement learning (RL), the question of learning policies that robustly satisfy state constraints under unknown disturbances remains open. In this paper, we offer a new perspective on achieving…

Machine Learning · Computer Science 2025-12-23 Pierre-François Massiani , Alexander von Rohr , Lukas Haverbeck , Sebastian Trimpe

We revisit the cut-off prescriptions which are needed in order to specify completely the form of Tsallis' maximum entropy distributions. For values of the Tsallis entropic parameter $q>1$ we advance an alternative cut-off prescription and…

Statistical Mechanics · Physics 2009-11-11 A. M. Teweldeberhan , A. R. Plastino , H. G. Miller

This paper investigates applicability of thermodynamic concepts and principles to competitive systems. We show that Tsallis entropies are suitable for characterisation of systems with transitive competition when mutations deviate from Gibbs…

Adaptation and Self-Organizing Systems · Physics 2014-03-10 A. Y. Klimenko

Tsallis and R\'{e}nyi entropy measures are two possible different generalizations of the Boltzmann-Gibbs entropy (or Shannon's information) but are not generalizations of each others. It is however the Sharma-Mittal measure, which was…

Statistical Mechanics · Physics 2014-10-13 Marco Masi

Entropy regularization is an important idea in reinforcement learning, with great success in recent algorithms like Soft Q Network (SQN) and Soft Actor-Critic (SAC1). In this work, we extend this idea into the on-policy realm. We propose…

Machine Learning · Computer Science 2020-10-19 Jingbin Liu , Xinyang Gu , Shuai Liu

The entropy of a closure operator has been recently proposed for the study of network coding and secret sharing. In this paper, we study closure operators in relation to their entropy. We first introduce four different kinds of rank…

Information Theory · Computer Science 2013-07-24 Maximilien Gadouleau

We present a technique for entropy optimization to calculate a distribution from its moments. The technique is based upon maximizing a discretized form of the Shannon entropy functional by mapping the problem onto a dual space where an…

Disordered Systems and Neural Networks · Physics 2009-11-10 K. Bandyopadhyay , A. K. Bhattacharya , Parthapratim Biswas , D. A. Drabold

Entropy is a key measure in studies related to information theory and its many applications. Campbell of the first time recognized that exponential of Shannons entropy is just the size of the sample space when the distribution is uniform.…

Information Theory · Computer Science 2016-08-15 Dhanesh Garg , Staish Kumar

The exact solution of a particular form of the stationary state generalized Fokker-Planck equations, which is given under certain conditions by the classical Tsallis distribution, is compared with the solution of the MAXENT equations…

Statistical Mechanics · Physics 2013-02-01 J. M. Conroy , H. G. Miller

State entropy regularization has empirically shown better exploration and sample complexity in reinforcement learning (RL). However, its theoretical guarantees have not been studied. In this paper, we show that state entropy regularization…

Machine Learning · Computer Science 2025-12-02 Yonatan Ashlag , Uri Koren , Mirco Mutti , Esther Derman , Pierre-Luc Bacon , Shie Mannor

Gauss' law of error is generalized in Tsallis statistics such as multifractal systems, in which Tsallis entropy plays an essential role instead of Shannon entropy. For the generalization, we apply the new multiplication operation determined…

Statistical Mechanics · Physics 2007-05-23 Hiroki Suyari , Makoto Tsukada

It is possible to derive the maximum entropy principle from thermodynamic stability requirements. Using as a starting point the equilibrium probability distribution, currently used in non-extensive thermostatistics, it turns out that the…

Statistical Mechanics · Physics 2007-05-23 Jan Naudts

The generalisation and robustness properties of policies learnt through Maximum-Entropy Reinforcement Learning are investigated on chaotic dynamical systems with Gaussian noise on the observable. First, the robustness under noise…

Machine Learning · Computer Science 2026-02-25 Rémy Hosseinkhan-Boucher , Onofrio Semeraro , Lionel Mathelin

We study the nonextensive thermodynamics for open systems. On the basis of the maximum entropy principle, the dual power-law q-distribution functions are re-deduced by using the dual particle number definitions and assuming that the…

Statistical Mechanics · Physics 2020-02-26 Yahui Zheng , Haining Yu , Jiulin Du

This work uses the entropy-regularised relaxed stochastic control perspective as a principled framework for designing reinforcement learning (RL) algorithms. Herein agent interacts with the environment by generating noisy controls…

Machine Learning · Computer Science 2023-09-18 Lukasz Szpruch , Tanut Treetanthiploet , Yufei Zhang

Stochastic and soft optimal policies resulting from entropy-regularized Markov decision processes (ER-MDP) are desirable for exploration and imitation learning applications. Motivated by the fact that such policies are sensitive with…

Machine Learning · Computer Science 2022-01-03 Tien Mai , Patrick Jaillet

Entropy regularization is an efficient technique for encouraging exploration and preventing a premature convergence of (vanilla) policy gradient methods in reinforcement learning (RL). However, the theoretical understanding of…

Machine Learning · Computer Science 2024-07-16 Yuhao Ding , Junzi Zhang , Hyunin Lee , Javad Lavaei

The paper extends the analysis of the entropies of the Poisson distribution with parameter $\lambda$. It demonstrates that the Tsallis and Sharma-Mittal entropies exhibit monotonic behavior with respect to $\lambda$, whereas two generalized…

Probability · Mathematics 2024-11-27 Dmitri Finkelshtein , Anatoliy Malyarenko , Yuliya Mishura , Kostiantyn Ralchenko

In this work, we present a novel characterization of approximate Nash equilibria in a class of convex games over the simplex. To achieve this, we regularize the utility functions using the Shannon entropy term, connect the solutions to the…

Optimization and Control · Mathematics 2025-07-18 Tatiana Tatarenko , S. Rasoul Etesami

Reinforcement learning (RL) has become a key approach for enhancing reasoning in large language models (LLMs), yet scalable training is often hindered by the rapid collapse of policy entropy, which leads to premature convergence and…

Machine Learning · Computer Science 2026-04-14 Ming Lei , Christophe Baehr
‹ Prev 1 4 5 6 7 8 10 Next ›