English
Related papers

Related papers: Reinforcement learning approach to non-equilibrium…

200 papers

Reinforcement learning has traditionally been studied with exponential discounting or the average reward setup, mainly due to their mathematical tractability. However, such frameworks fall short of accurately capturing human behavior, which…

Machine Learning · Computer Science 2024-09-18 S. R. Eshwar , Mayank Motwani , Nibedita Roy , Gugan Thoppe

This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type of exploratory formulation under entropy regularization where the agent randomizes both the timing…

Optimization and Control · Mathematics 2025-12-23 Yijie Huang , Mengge Li , Xiang Yu , Zhou Zhou

We define and analyse the concept of entanglement production during the evolution of a general quantum mechanical dissipative system. While it is important to minimise entropy production in order to achieve thermodynamical efficiency,…

Quantum Physics · Physics 2007-06-22 V. Vedral

In practical applications, we can rarely assume full observability of a system's environment, despite such knowledge being important for determining a reactive control system's precise interaction with its environment. Therefore, we propose…

Machine Learning · Computer Science 2022-06-24 Edi Muskardin , Martin Tappler , Bernhard K. Aichernig , Ingo Pill

Determining the Hamiltonian of a quantum system is essential for understanding its dynamics and validating its behavior. Hamiltonian learning provides a data-driven approach to reconstruct the generator of the dynamics from measurements on…

As the number of qubits in a sensor increases, the complexity of designing and controlling the quantum circuits grows exponentially. Manually optimizing these circuits becomes infeasible. Optimizing entanglement distribution in large-scale…

Quantum Physics · Physics 2025-09-01 Laxmisha Ashok Attisara , Sathish Kumar

We show explicitly the entropy reduction from a detailed fluctuation theorem for the general stochastic system driven by nonequilibrium process under feedback control. The effect of interaction of the feedback controller with the system is…

Statistical Mechanics · Physics 2011-04-27 M. Ponmurugan

Quantum control has been of increasing interest in recent years, e.g. for tasks like state initialization and stabilization. Feedback-based strategies are particularly powerful, but also hard to find, due to the exponentially increased…

Quantum Physics · Physics 2022-06-30 Riccardo Porotti , Antoine Essig , Benjamin Huard , Florian Marquardt

A reinforcement algorithm solves a classical optimization problem by introducing a feedback to the system which slowly changes the energy landscape and converges the algorithm to an optimal solution in the configuration space. Here, we use…

Disordered Systems and Neural Networks · Physics 2017-11-08 A. Ramezanpour

Traditionally, reinforcement learning methods predict the next action based on the current state. However, in many situations, directly applying actions to control systems or robots is dangerous and may lead to unexpected behaviors because…

Robotics · Computer Science 2020-11-03 Nan Lin , Yuxuan Li , Yujun Zhu , Ruolin Wang , Xiayu Zhang , Jianmin Ji , Keke Tang , Xiaoping Chen , Xinming Zhang

Batch reinforcement learning enables policy learning without direct interaction with the environment during training, relying exclusively on previously collected sets of interactions. This approach is, therefore, well-suited for high-risk…

Machine Learning · Computer Science 2024-11-18 Amna Najib , Stefan Depeweg , Phillip Swazinna

We set up a framework for quantum stochastic thermodynamics based solely on experimentally controllable, but otherwise arbitrary interventions at discrete times. Using standard assumptions about the system-bath dynamics and insights from…

Quantum Physics · Physics 2019-08-22 Philipp Strasberg

This paper presents a constrained policy gradient algorithm. We introduce constraints for safe learning with the following steps. First, learning is slowed down (lazy learning) so that the episodic policy change can be computed with the…

Machine Learning · Computer Science 2022-01-24 Balázs Varga , Balázs Kulcsár , Morteza Haghir Chehreghani

A data-efficient learning-based control design method is proposed in this paper. It is based on learning a system dynamics model that is then leveraged in a two-level procedure. On the higher level, a simple but powerful optimization…

Systems and Control · Electrical Eng. & Systems 2026-02-03 Ludvig Svedlund , Constantin Cronrath , Jonas Fredriksson , Bengt Lennartson

We propose a comprehensive framework for policy gradient methods tailored to continuous time reinforcement learning. This is based on the connection between stochastic control problems and randomised problems, enabling applications across…

Optimization and Control · Mathematics 2024-05-01 Robert Denkert , Huyên Pham , Xavier Warin

Experimental advances enabling high-resolution external control create new opportunities to produce materials with exotic properties. In this work, we investigate how a multi-agent reinforcement learning approach can be used to design…

Statistical Mechanics · Physics 2021-11-15 Shriram Chennakesavalu , Grant M. Rotskoff

Electric water heaters have the ability to store energy in their water buffer without impacting the comfort of the end user. This feature makes them a prime candidate for residential demand response. However, the stochastic and nonlinear…

Machine Learning · Computer Science 2015-12-02 Frederik Ruelens , Bert Claessens , Salman Quaiyum , Bart De Schutter , Robert Babuska , Ronnie Belmans

Reinforcement learning is a powerful paradigm for learning optimal policies from experimental data. However, to find optimal policies, most reinforcement learning algorithms explore all possible actions, which may be harmful for real-world…

Machine Learning · Statistics 2017-11-15 Felix Berkenkamp , Matteo Turchetta , Angela P. Schoellig , Andreas Krause

Informational contributions to thermodynamics can be studied in isolation by considering systems with fully-degenerate Hamiltonians. In this regime, being in non-equilibrium -- termed informational non-equilibrium -- provides thermodynamic…

Quantum Physics · Physics 2025-05-15 Chung-Yun Hsieh , Benjamin Stratton , Hao-Cheng Weng , Valerio Scarani

We explore the reinforcement learning approach to designing controllers by extensively discussing the case of a quadcopter attitude controller. We provide all details allowing to reproduce our approach, starting with a model of the dynamics…

Artificial Intelligence · Computer Science 2021-07-28 Nicola Bernini , Mikhail Bessa , Rémi Delmas , Arthur Gold , Eric Goubault , Romain Pennec , Sylvie Putot , François Sillion