English
Related papers

Related papers: Generalized Maximum Entropy Differential Dynamic P…

200 papers

We propose a method to derive the stationary size distributions of a system, and the degree distributions of networks, using maximisation of the Gibbs-Shannon entropy. We apply this to a preferential attachment-type algorithm for systems of…

Physics and Society · Physics 2020-03-17 Cornelia Metzig , Caroline Colijn

This paper shows how to evolve numerically the maximum entropy probability distributions for a given set of constraints, which is a variational calculus problem. An evolutionary algorithm can obtain approximations to some well-known…

Methodology · Statistics 2020-02-07 Raul Rojas

Recently deep reinforcement learning (DRL) has achieved outstanding success on solving many difficult and large-scale RL problems. However the high sample cost required for effective learning often makes DRL unaffordable in resource-limited…

Machine Learning · Computer Science 2018-09-06 Gang Chen , Yiming Peng , Mengjie Zhang

We address the challenge of exploration in reinforcement learning (RL) when the agent operates in an unknown environment with sparse or no rewards. In this work, we study the maximum entropy exploration problem of two different types. The…

We investigate the cumulative Tsallis entropy, an information measure recently introduced as a cumulative version of the classical Tsallis differential entropy, which is itself a generalization of the Boltzmann-Gibbs statistics. This…

Statistics Theory · Mathematics 2023-06-02 Guillaume Dulac , Thomas Simon

In this paper, we propose a novel maximum causal Tsallis entropy (MCTE) framework for imitation learning which can efficiently learn a sparse multi-modal policy distribution from demonstrations. We provide the full mathematical analysis of…

Machine Learning · Computer Science 2018-05-29 Kyungjae Lee , Sungjoon Choi , Songhwai Oh

Shannon entropy regularization is widely adopted in optimal control due to its ability to promote exploration and enhance robustness, e.g., maximum entropy reinforcement learning known as Soft Actor-Critic. In this paper, Tsallis entropy,…

Optimization and Control · Mathematics 2024-03-05 Yota Hashizume , Koshi Oishi , Kenji Kashima

This paper investigates an optimal consumption-investment problem featuring recursive utility via Tsallis relative entropy. We establish a fundamental connection between this optimization problem and a quadratic backward stochastic…

Mathematical Finance · Quantitative Finance 2025-09-26 Xueying Huang , Peng Luo , Dejian Tian

Generalized evolutionary algorithm based on Tsallis canonical distribution is proposed. The algorithm uses Tsallis generalized canonical distribution to weigh the configurations for `selection' instead of Gibbs-Boltzmann distribution. Our…

Artificial Intelligence · Computer Science 2007-05-23 Ambedkar Dukkipati , M. Narasimha Murty , Shalabh Bhatnagar

Many Imitation and Reinforcement Learning approaches rely on the availability of expert-generated demonstrations for learning policies or value functions from data. Obtaining a reliable distribution of trajectories from motion planners is…

Robotics · Computer Science 2021-07-13 Alexander Lambert , Byron Boots

Trajectory optimization methods for motion planning attempt to generate trajectories that minimize a suitable objective function. Such methods efficiently find solutions even for high degree-of-freedom robots. However, a globally optimal…

Robotics · Computer Science 2019-07-18 Luka Petrović , Juraj Peršić , Marija Seder , Ivan Marković

We study the evolution of Tsallis entropy along the heat flow and establish its concavity in arbitrary dimensions. Extending prior results that were restricted to the one-dimensional setting, we prove that the Tsallis entropy is concave in…

Information Theory · Computer Science 2026-04-24 Lukang Sun

Bayesian optimization through Gaussian process regression is an effective method of optimizing an unknown function for which every measurement is expensive. It approximates the objective function and then recommends a new measurement point…

Machine Learning · Statistics 2017-05-17 Hildo Bijl , Thomas B. Schön , Jan-Willem van Wingerden , Michel Verhaegen

We revisit the well-studied problem of estimating the Shannon entropy of a probability distribution, now given access to a probability-revealing conditional sampling oracle. In this model, the oracle takes as input the representation of a…

Cryptography and Security · Computer Science 2022-06-03 Priyanka Golia , Brendan Juba , Kuldeep S. Meel

Trajectory optimization and model predictive control are essential techniques underpinning advanced robotic applications, ranging from autonomous driving to full-body humanoid control. State-of-the-art algorithms have focused on data-driven…

Systems and Control · Electrical Eng. & Systems 2021-11-15 Hany Abdulsamad , Tim Dorau , Boris Belousov , Jia-Jie Zhu , Jan Peters

We propose a Gaussian variational inference framework for the motion planning problem. In this framework, motion planning is formulated as an optimization over the distribution of the trajectories to approximate the desired trajectory…

Robotics · Computer Science 2023-03-27 Hongzhe Yu , Yongxin Chen

We give a new proof of the theorems on the maximum entropy principle in Tsallis statistics. That is, we show that the $q$-canonical distribution attains the maximum value of the Tsallis entropy, subject to the constraint on the…

Statistical Mechanics · Physics 2015-05-14 Shigeru Furuichi

We study the sparse entropy-regularized reinforcement learning (ERL) problem in which the entropy term is a special form of the Tsallis entropy. The optimal policy of this formulation is sparse, i.e.,~at each state, it has non-zero…

Artificial Intelligence · Computer Science 2018-02-13 Ofir Nachum , Yinlam Chow , Mohammad Ghavamzadeh

This paper addresses the problem of dynamic asset allocation under uncertainty, which can be formulated as a linear quadratic (LQ) control problem with multiplicative noise. To handle exploration exploitation trade offs and induce sparse…

Optimization and Control · Mathematics 2025-09-30 Haoran Zhang , Wenhao Zhang , Xianping Wu

We propose a generalized entropy maximization procedure, which takes into account the generalized averaging procedures and information gain definitions underlying the generalized entropies. This novel generalized procedure is then applied…

Statistical Mechanics · Physics 2015-05-13 G. Baris Bagci , Ugur Tirnakli