中文
相关论文

相关论文: A Demon that remembers: An agential approach towar…

200 篇论文

In the last decade quantum machine learning has provided fascinating and fundamental improvements to supervised, unsupervised and reinforcement learning. In reinforcement learning, a so-called agent is challenged to solve a task given by…

量子物理 · 物理学 2022-04-13 Arne Hamann , Sabine Wölk

We investigate the limitations that emerge in thermodynamic tasks as a result of having local control only over the components of a thermal machine. These limitations are particularly relevant for devices composed of interacting many-body…

量子物理 · 物理学 2018-03-07 J. Lekscha , H. Wilming , J. Eisert , R. Gallego

LLM agents increasingly operate in open-ended environments spanning hundreds of sequential episodes, yet they remain largely stateless: each task is solved from scratch without converting past experience into better future behavior. The…

计算与语言 · 计算机科学 2026-04-24 Wujiang Xu , Jiaojiao Han , Minghao Guo , Kai Mei , Xi Zhu , Han Zhang , Dimitris N. Metaxas

Dissipative adaptation is a general thermodynamic mechanism that explains self-organization in a broad class of driven classical many-body systems. It establishes how the most likely (adapted) states of a system subjected to a given drive…

量子物理 · 物理学 2021-11-17 Daniel Valente , Frederico Brito , Thiago Werlang

In this study, we introduce a novel approach in quantum field theories to estimate the action using the artificial neural networks (ANNs). The estimation is achieved by learning on system configurations governed by the Boltzmann factor,…

高能物理 - 格点 · 物理学 2024-10-10 Tian Xu , Lingxiao Wang , Lianyi He , Kai Zhou , Yin Jiang

Experience-driven learning has emerged as a promising paradigm for enabling agents to improve from interaction trajectories by accumulating and reusing past experience. However, existing approaches are predominantly developed in textual…

人工智能 · 计算机科学 2026-05-19 Xingyu Sui , Weixiang Zhao , Yongxin Tang , Yanyan Zhao , Yang Wu , Dandan Tu , Bing Qin

Motivated by the recent interest in thermodynamics of micro- and mesoscopic quantum systems we study the maximal amount of work that can be reversibly extracted from a quantum system used to store temporarily energy. Guided by the notion of…

量子物理 · 物理学 2013-05-01 Robert Alicki , Mark Fannes

We design a heat engine with multi-heat-reservoir, ancillary system and quantum memory. We then derive an inequality related with the second law of thermodynamics, and give a new limitation about the work gain from the engine by analyzing…

量子物理 · 物理学 2017-11-09 Li-Hang Ren , Heng Fan

Cooperation emergence in multi-agent systems represents a fundamental statistical physics problem where microscopic learning rules drive macroscopic collective behavior transitions. We propose a Q-learning-based variant of adaptive rewiring…

物理与社会 · 物理学 2025-09-04 Yi-Ning Weng , Hsuan-Wei Lee

Evaluating the maximum amount of work extractable from a nanoscale quantum system is one of the central problems in quantum thermodynamics. Previous works identified the free energy of the input state as the optimal rate of extractable work…

量子物理 · 物理学 2026-03-06 Kaito Watanabe , Ryuji Takagi

We develop a resource-theoretic framework that allows one to bridge the gap between two approaches to quantum thermodynamics based on Markovian thermal processes (which model memoryless dynamics) and thermal operations (which model…

量子物理 · 物理学 2023-10-10 Jakub Czartowski , A. de Oliveira Junior , Kamil Korzekwa

Modeling the complex interactions of systems of particles or agents is a fundamental scientific and mathematical problem that is studied in diverse fields, ranging from physics and biology, to economics and machine learning. In this work,…

机器学习 · 统计学 2020-10-09 Jason Miller , Sui Tang , Ming Zhong , Mauro Maggioni

We ground the asymmetry of causal relations in the internal physical states of a special kind of open and irreversible physical system, a causal agent. A causal agent is an autonomous physical system, maintained in a steady state, far from…

物理学史与哲学 · 物理学 2023-06-05 G. J. Milburn , S. Shrapnel , P. W. Evans

There is a deep connection between thermodynamics, information and work extraction. Ever since the birth of thermodynamics, various types of Maxwell demons have been introduced in order to deepen our understanding of the second law. Thanks…

统计力学 · 物理学 2024-08-14 Francesco Caravelli

Active inference is a formal approach to study cognition based on the notion that adaptive agents can be seen as engaging in a process of approximate Bayesian inference, via the minimisation of variational and expected free energies.…

人工智能 · 计算机科学 2025-08-19 Filippo Torresan , Keisuke Suzuki , Ryota Kanai , Manuel Baltieri

We use a reinforcement learning approach to reduce entropy production in a closed quantum system brought out of equilibrium. Our strategy makes use of an external control Hamiltonian and a policy gradient technique. Our approach bears no…

量子物理 · 物理学 2024-02-21 Sofia Sgroi , G. Massimo Palma , Mauro Paternostro

Robotic manipulation often requires memory: occlusion and state changes can make decision-time observations perceptually aliased, making action selection non-Markovian at the observation level because the same observation may arise from…

机器人学 · 计算机科学 2026-03-26 Xinying Guo , Chenxi Jiang , Hyun Bin Kim , Ying Sun , Yang Xiao , Yuhang Han , Jianfei Yang

In model-based reinforcement learning, generative and temporal models of environments can be leveraged to boost agent performance, either by tuning the agent's representations during training or via use as part of an explicit planning…

Nowadays, model-free reinforcement learning algorithms have achieved remarkable performance on many decision making and control tasks, but high sample complexity and low sample efficiency still hinder the wide use of model-free…

人工智能 · 计算机科学 2020-10-27 Jingbin Liu , Xinyang Gu , Shuai Liu

Inspired by recent work in attention models for image captioning and question answering, we present a soft attention model for the reinforcement learning domain. This model uses a soft, top-down attention mechanism to create a bottleneck in…

机器学习 · 计算机科学 2019-06-07 Alex Mott , Daniel Zoran , Mike Chrzanowski , Daan Wierstra , Danilo J. Rezende