English
Related papers

Related papers: Exploratory Control with Tsallis Entropy for Laten…

200 papers

In this paper, we consider a discrete-time stochastic control problem with uncertain initial and target states. We first discuss the connection between optimal transport and stochastic control problems of this form. Next, we formulate a…

Maximum entropy reinforcement learning motivates agents to explore states and actions to maximize the entropy of some distribution, typically by providing additional intrinsic rewards proportional to that entropy function. In this paper, we…

Machine Learning · Computer Science 2026-03-20 Adrien Bolland , Gaspard Lambrechts , Damien Ernst

Sample-based trajectory optimisers are a promising tool for the control of robotics with non-differentiable dynamics and cost functions. Contemporary approaches derive from a restricted subclass of stochastic optimal control where the…

Robotics · Computer Science 2021-10-07 Tom Lefebvre , Guillaume Crevecoeur

We investigate the cumulative Tsallis entropy, an information measure recently introduced as a cumulative version of the classical Tsallis differential entropy, which is itself a generalization of the Boltzmann-Gibbs statistics. This…

Statistics Theory · Mathematics 2023-06-02 Guillaume Dulac , Thomas Simon

Optimal control of stochastic nonlinear dynamical systems is a major challenge in the domain of robot learning. Given the intractability of the global control problem, state-of-the-art algorithms focus on approximate sequential optimization…

Machine Learning · Computer Science 2020-04-23 Joe Watson , Hany Abdulsamad , Jan Peters

Current research on robust trajectory planning for autonomous agents aims to mitigate uncertainties arising from disturbances and modeling errors while ensuring guaranteed safety. Existing methods primarily utilize stochastic optimal…

Systems and Control · Electrical Eng. & Systems 2025-02-13 Christian Vitale , Savvas Papaioannou , Panayiotis Kolios , Georgios Ellinas

This paper investigates the optimal control problem for a class of nonlinear fully coupled forward-backward stochastic difference equations (FBS$\Delta$Es). Under the convexity assumption of the control domain, we establish a variational…

Optimization and Control · Mathematics 2025-12-02 Zhipeng Niu , Jun Moon , Qingxin Meng

In a reinforcement learning (RL) framework, we study the exploratory version of the continuous time expected utility (EU) maximization problem with a portfolio constraint that includes widely-used financial regulations such as short-selling…

Mathematical Finance · Quantitative Finance 2024-12-17 Huy Chau , Duy Nguyen , Thai Nguyen

We study the discrete-time linear-quadratic (LQ) control model using reinforcement learning (RL). Using entropy to measure the cost of exploration, we prove that the optimal feedback policy for the problem must be Gaussian type. Then, we…

Machine Learning · Statistics 2025-02-05 Lucky Li

This paper is concerned with the linear quadratic (LQ) optimal control of continuous-time system with terminal state constraint. In particular, multiple agents exist in the system which can only access partial information of the matrix…

Optimization and Control · Mathematics 2025-10-21 Wenjing Yang , Zhaorong Zhang , Juanjuan Xu

The Tsallis entropy, which is a generalization of the Boltzmann-Gibbs entropy, plays a central role in nonextensive statistical mechanics of complex systems. A lot of efforts have recently been made on establishing a dynamical foundation…

Statistical Mechanics · Physics 2009-11-11 Sumiyoshi Abe , Yutaka Nakada

The Soft Actor-Critic (SAC) algorithm with a Gaussian policy has become a mainstream implementation for realizing the Maximum Entropy Reinforcement Learning (MaxEnt RL) objective, which incorporates entropy maximization to encourage…

Machine Learning · Computer Science 2025-06-09 Xiaoyi Dong , Jian Cheng , Xi Sheryl Zhang

We study a general class of entropy-regularized multi-variate LQG mean field games (MFGs) in continuous time with $K$ distinct sub-population of agents. We extend the notion of actions to action distributions (exploratory actions), and…

Optimization and Control · Mathematics 2021-12-01 Dena Firoozi , Sebastian Jaimungal

This paper is devoted to an optimal control problem of fully coupled forward-backward stochastic differential equations driven by sub-diffusion, whose solutions are not Markov processes. The stochastic maximum principle is obtained, where…

Optimization and Control · Mathematics 2025-03-11 Chenhui Hao , Jingtao Shi , Shuaiqi Zhang

In this paper, we investigate new procedures for statistical testing based on Tsallis entropy, a parametric generalization of Shannon entropy. Focusing on multivariate generalized Gaussian and $q$-Gaussian distributions, we develop…

Methodology · Statistics 2025-06-18 Mehmet Sıddık Çadırcı

We prove a general existence result in stochastic optimal control in discrete time where controls take values in conditional metric spaces, and depend on the current state and the information of past decisions through the evolution of a…

Optimization and Control · Mathematics 2018-12-19 Asgar Jamneshan , Michael Kupper , José Miguel Zapata

We demonstrate that the most probable state of a conserved system with a limited number of entities or molecules is the state where non-Gaussian and non-chi-square distributions govern. We have conducted a thought experiment using a…

Mathematical Physics · Physics 2023-04-20 Jae Wan Shim

We investigate the optimal control of large-scale autonomous systems under explicitly adversarial conditions, incorporating the probabilistic destruction of agents over time. In many such systems, adversarial interactions arise as different…

Optimization and Control · Mathematics 2026-02-27 Claire Walton , Isaac Kaminer , Qi Gong , Abram H. Clark , Theodoros Tsatsanifos

We first observe that the (co)domains of the q-deformed functions are some subsets of the (co)domains of their ordinary counterparts, thereby deeming the deformed functions to be incomplete. In order to obtain a complete definition of…

Statistical Mechanics · Physics 2015-05-13 Thomas Oikonomou , G. Baris Bagci

We consider the problem of controlling the group behavior of a large number of dynamic systems that are constantly interacting with each other. These systems are assumed to have identical dynamics (e.g., birds flock, robot swarm) and their…

Optimization and Control · Mathematics 2021-08-18 Yongxin Chen