English
Related papers

Related papers: Tsallis Entropy Regularization for Linearly Solvab…

200 papers

In this paper, we present a new class of Markov decision processes (MDPs), called Tsallis MDPs, with Tsallis entropy maximization, which generalizes existing maximum entropy reinforcement learning (RL). A Tsallis MDP provides a unified…

Machine Learning · Computer Science 2019-02-08 Kyungjae Lee , Sungyub Kim , Sungbin Lim , Sungjoon Choi , Songhwai Oh

This paper addresses the problem of dynamic asset allocation under uncertainty, which can be formulated as a linear quadratic (LQ) control problem with multiplicative noise. To handle exploration exploitation trade offs and induce sparse…

Optimization and Control · Mathematics 2025-09-30 Haoran Zhang , Wenhao Zhang , Xianping Wu

Maximum Tsallis entropy (MTE) framework in reinforcement learning has gained popularity recently by virtue of its flexible modeling choices including the widely used Shannon entropy and sparse entropy. However, non-Shannon entropies suffer…

Machine Learning · Computer Science 2022-05-18 Lingwei Zhu , Zheng Chen , Eiji Uchibe , Takamitsu Matsubara

We study the sparse entropy-regularized reinforcement learning (ERL) problem in which the entropy term is a special form of the Tsallis entropy. The optimal policy of this formulation is sparse, i.e.,~at each state, it has non-zero…

Artificial Intelligence · Computer Science 2018-02-13 Ofir Nachum , Yinlam Chow , Mohammad Ghavamzadeh

This paper studies the continuous-time reinforcement learning in jump-diffusion models by featuring the q-learning (the continuous-time counterpart of Q-learning) under Tsallis entropy regularization. Contrary to the Shannon entropy, the…

Optimization and Control · Mathematics 2026-02-16 Lijun Bo , Yijie Huang , Xiang Yu , Tingting Zhang

The Tsallis entropy given for a positive parameter $\alpha$ can be considered as a modification of the classical Shannon entropy. For the latter, corresponding to $\alpha=1$, there exist many axiomatic characterizations. One of them based…

Mathematical Physics · Physics 2017-04-27 Sonja Jäckle , Karsten Keller

In this paper, a sparse Markov decision process (MDP) with novel causal sparse Tsallis entropy regularization is proposed.The proposed policy regularization induces a sparse and multi-modal optimal policy distribution of a sparse MDP. The…

Machine Learning · Computer Science 2017-10-16 Kyungjae Lee , Sungjoon Choi , Songhwai Oh

Within a framework of utmost generality, we show that the entropy maximization procedure with linear constraints uniquely leads to the Shannon-Boltzmann-Gibbs entropy. Therefore, the use of this procedure with linear constraints should not…

Statistical Mechanics · Physics 2018-05-01 Thomas Oikonomou , G. Baris Bagci

The purpose of this note is to give the general solution of two functional equations connected to the Shannon entropy and also to the Tsallis entropy. As a result of this, we present the regular solution of these equations, as well.…

Classical Analysis and ODEs · Mathematics 2013-07-03 Eszter Gselmann

Tsallis relative operator entropy was defined as a parametric extension of relative operator entropy and the generalized Shannon inequalities were shown in the previous paper. After the review of some fundamental properties of Tsallis…

Functional Analysis · Mathematics 2010-01-10 Shigeru Furuichi , Kenjiro Yanagi , Ken Kuriyama

Policy regularization methods such as maximum entropy regularization are widely used in reinforcement learning to improve the robustness of a learned policy. In this paper, we show how this robustness arises from hedging against worst-case…

Machine Learning · Computer Science 2024-04-29 Rob Brekelmans , Tim Genewein , Jordi Grau-Moya , Grégoire Delétang , Markus Kunesch , Shane Legg , Pedro Ortega

We study optimal control in models with latent factors where the agent controls the distribution over actions, rather than actions themselves, in both discrete and continuous time. To encourage exploration of the state space, we reward…

Mathematical Finance · Quantitative Finance 2024-01-03 Ryan Donnelly , Sebastian Jaimungal

By using the maximum entropy principle with Tsallis entropy we obtain a fragment size distribution function which undergoes a transition to scaling. This distribution function reduces to those obtained by other authors using Shannon…

Soft Condensed Matter · Physics 2015-06-24 Oscar Sotolongo-Costa , Arezky H. Rodriguez , G. J. Rodgers

Recently deep reinforcement learning (DRL) has achieved outstanding success on solving many difficult and large-scale RL problems. However the high sample cost required for effective learning often makes DRL unaffordable in resource-limited…

Machine Learning · Computer Science 2018-09-06 Gang Chen , Yiming Peng , Mengjie Zhang

In this paper, we consider the information content of maximum ranked set sampling procedure with unequal samples (MRSSU) in terms of Tsallis entropy which is a nonadditive generalization of Shannon entropy. We obtain several results of…

Statistics Theory · Mathematics 2020-11-04 S. Tahmasebi , M. Longobardi , M. R. Kazemi , M. Alizadeh

The q-exponential distributions, which are generalizations of the Zipf-Mandelbrot power-law distribution, are frequently encountered in complex systems at their stationary states. From the viewpoint of the principle of maximum entropy, they…

Statistical Mechanics · Physics 2009-11-07 Sumiyoshi Abe

We present a sampling-based trajectory optimization method derived from the maximum entropy formulation of Differential Dynamic Programming with Tsallis entropy. This method is a generalization of the legacy work with Shannon entropy, which…

Optimization and Control · Mathematics 2024-09-18 Yuichiro Aoyama , Evangelos A. Theodorou

In recent years, learning for neural networks can be viewed as optimization in the space of probability measures. To obtain the exponential convergence to the optimizer, the regularizing term based on Shannon entropy plays an important…

Machine Learning · Statistics 2024-11-07 Keito Akiyama

This article extends the non-extensive entropy of Tsallis and uses this entropy to model an energy producing system in an absorbing heat bath. This modified non-extensive entropy is superficially identical to the one proposed by Tsallis,…

Statistical Mechanics · Physics 2007-05-23 Mark Fleischer

In this research paper, it is proved that an approximation to Gibbs-Shannon entropy measure naturally leads to Tsallis entropy for the real parameter q =2 . Several interesting measures based on the input as well as output of a discrete…

Information Theory · Computer Science 2012-01-06 Garimella Rama Murthy
‹ Prev 1 2 3 10 Next ›