English
Related papers

Related papers: Data-Driven Long-Term Asset Allocation with Tsalli…

200 papers

We introduce a generic solver for dynamic portfolio allocation problems when the market exhibits return predictability, price impact and partial observability. We assume that the price modeling can be encoded into a linear state-space and…

Portfolio Management · Quantitative Finance 2016-11-07 M. Abeille , E. Serie , A. Lazaric , X. Brokmann

In this paper we consider the dynamic Tsallis entropy and employ it for four model systems: (i) the motion of Brownian oscillator, (ii) the motion of Brownian oscillator with noise, (iii) the fluctuation of particle density in hydrodynamics…

Other Condensed Matter · Physics 2007-05-23 Nail R. Khusnutdinov , Renat M. Yulmetyev , Natalya A. Emelyanova

This paper proposes a novel approach for Asset-Liability Management (ALM) by employing continuous-time Reinforcement Learning (RL) with a linear-quadratic (LQ) formulation that incorporates both interim and terminal objectives. We develop a…

Machine Learning · Computer Science 2025-09-30 Yilie Huang

Entropy regularization is used to get improved optimization performance in reinforcement learning tasks. A common form of regularization is to maximize policy entropy to avoid premature convergence and lead to more stochastic policies for…

Machine Learning · Computer Science 2019-12-12 Riashat Islam , Zafarali Ahmed , Doina Precup

We study the sparse entropy-regularized reinforcement learning (ERL) problem in which the entropy term is a special form of the Tsallis entropy. The optimal policy of this formulation is sparse, i.e.,~at each state, it has non-zero…

Artificial Intelligence · Computer Science 2018-02-13 Ofir Nachum , Yinlam Chow , Mohammad Ghavamzadeh

We consider dynamical systems evolving near an equilibrium statistical state where the interest is in modelling long term behavior that is consistent with thermodynamic constraints. We adjust the distribution using an entropy-optimizing…

Fluid Dynamics · Physics 2014-11-25 Keith Myerscough , Jason Frank , Benedict Leimkuhler

A finite horizon linear quadratic(LQ) optimal control problem is studied for a class of discrete-time linear fractional systems (LFSs) affected by multiplicative, independent random perturbations. Based on the dynamic programming technique,…

Optimization and Control · Mathematics 2016-07-01 J. J. Trujillo , V. M. Ungureanu

Traditional thermodynamic trade-off relations usually apply to quantities that depend linearly on probability distributions. In contrast, many important information-theoretic measures, such as entropies, are nonlinear and therefore…

Statistical Mechanics · Physics 2026-02-17 Yoshihiko Hasegawa

This paper bridges reinforcement learning (RL) and risk-sensitive stochastic control by introducing a tractable exploration mechanism for policy search in risk-sensitive portfolio management, with known and unknown model parameters, that…

Portfolio Management · Quantitative Finance 2026-03-03 Sebastien Lleo , Wolfgang Runggaldier

The linear quadratic regulator (LQR) problem is a cornerstone of automatic control, and it has been widely studied in the data-driven setting. The various data-driven approaches can be classified as indirect (i.e., based on an identified…

Optimization and Control · Mathematics 2021-09-15 Florian Dörfler , Pietro Tesi , Claudio De Persis

Stabilization of linear systems with unknown dynamics is a canonical problem in adaptive control. Since the lack of knowledge of system parameters can cause it to become destabilized, an adaptive stabilization procedure is needed prior to…

Systems and Control · Computer Science 2018-07-25 Mohamad Kazem Shirani Faradonbeh , Ambuj Tewari , George Michailidis

Most data-driven analysis and control methods rely on centralized access to system measurements. In contrast, we consider a setting in which the measurements are distributed across multiple agents and raw data are not shared. Each agent has…

Optimization and Control · Mathematics 2026-03-12 Surya Malladi , Nima Monshizadeh

In this paper, we present a new class of Markov decision processes (MDPs), called Tsallis MDPs, with Tsallis entropy maximization, which generalizes existing maximum entropy reinforcement learning (RL). A Tsallis MDP provides a unified…

Machine Learning · Computer Science 2019-02-08 Kyungjae Lee , Sungyub Kim , Sungbin Lim , Sungjoon Choi , Songhwai Oh

We describe a simple and accurate framework for modeling the statistical behavior of both fully developed turbulence and short-term dynamics of financial markets based on the formalism of Tsallis' generalized non-extensive thermostatistics.…

Condensed Matter · Physics 2007-05-23 F. M. Ramos , C. Rodrigues Neto , R. R. Rosa

We address payoff-based decentralized learning in infinite-horizon zero-sum Markov games. In this setting, each player makes decisions based solely on received rewards, without observing the opponent's strategy or actions nor sharing…

Computer Science and Game Theory · Computer Science 2025-02-11 Reda Ouhamma , Maryam Kamgarpour

This paper presents a novel model-free and fully data-driven policy iteration scheme for quadratic regulation of linear dynamics with state- and input-multiplicative noise. The implementation is similar to the least-squares temporal…

Optimization and Control · Mathematics 2022-12-05 Peter Coppens , Panagiotis Patrinos

Stabilizing an unknown control system is one of the most fundamental problems in control systems engineering. In this paper, we provide a simple, model-free algorithm for stabilizing fully observed dynamical systems. While model-free…

Systems and Control · Electrical Eng. & Systems 2021-10-14 Juan C. Perdomo , Jack Umenberger , Max Simchowitz

This article investigates the core mechanisms of indirect data-driven control for unknown systems, focusing on the application of policy iteration (PI) within the context of the linear quadratic regulator (LQR) optimal control problem.…

Systems and Control · Electrical Eng. & Systems 2025-04-14 Bowen Song , Andrea Iannelli

Quite general, analytical (both exact and approximate) forms for discrete probability distributions (PD's) that maximize Tsallis entropy for a fixed variance are here investigated. They apply, for instance, in a wide variety of scenarios in…

Statistical Mechanics · Physics 2009-11-11 C. Vignat , A. Plastino

We propose a new risk-constrained formulation of the classical Linear Quadratic (LQ) stochastic control problem for general partially-observed systems. Our framework is motivated by the fact that the risk-neutral LQ controllers, although…

Optimization and Control · Mathematics 2021-12-15 Anastasios Tsiamis , Dionysios S. Kalogerias , Alejandro Ribeiro , George J. Pappas