English
Related papers

Related papers: HEAT:History-Enhanced Dual-phase Actor-Critic Algo…

200 papers

Hierarchies of temporally decoupled policies present a promising approach for enabling structured exploration in complex long-term planning problems. To fully achieve this approach an end-to-end training paradigm is needed. However,…

Machine Learning · Computer Science 2021-11-19 Abdul Rahman Kreidieh , Glen Berseth , Brandon Trabucco , Samyak Parajuli , Sergey Levine , Alexandre M. Bayen

Low-Power Wide Area Networks (LPWAN) play a key role in the IoT marketplace wherein LoRaWAN is considered a leading solution. Despite the traction of LoRaWAN, research shows that the current contention management mechanisms of LoRaWAN do…

Networking and Internet Architecture · Computer Science 2020-02-19 Stéphane Delbruel , Nicolas Small , Emekcan Aras , Jonathan Oostvogels , Danny Hughes

A VLAN is a logical connection that allows hosts to be grouped together in the same broadcast domain, so that packets are delivered only to ports that are combined to the same VLAN. We can improve wireless network performance and save…

Networking and Internet Architecture · Computer Science 2020-07-15 Tareq Al-Khraishi , Muhannad Quwaider

Model-based next state prediction and state value prediction are slow to converge. To address these challenges, we do the following: i) Instead of a neural network, we do model-based planning using a parallel memory retrieval system (which…

Artificial Intelligence · Computer Science 2023-02-02 John Chong Min Tan , Mehul Motani

Sequential decision-making agents struggle with long horizon tasks, since solving them requires multi-step reasoning. Most reinforcement learning (RL) algorithms address this challenge by improved credit assignment, introducing memory…

Machine Learning · Computer Science 2023-04-04 Bogdan Mazoure , Jake Bruce , Doina Precup , Rob Fergus , Ankit Anand

Iterative preference optimization methods have recently been shown to perform well for general instruction tuning tasks, but typically make little improvement on reasoning tasks (Yuan et al., 2024, Chen et al., 2024). In this work we…

Computation and Language · Computer Science 2024-06-27 Richard Yuanzhe Pang , Weizhe Yuan , Kyunghyun Cho , He He , Sainbayar Sukhbaatar , Jason Weston

Due to their expressive capacity, diffusion models have shown great promise in offline RL and imitation learning. Diffusion Actor-Critic with Entropy Regulator (DACER) extended this capability to online RL by using the reverse diffusion…

We propose a quantum heat transformer (QHT), a quantum thermodynamic device that modulates temperature gradients between two thermal junctions in quantum systems. Functionally, the QHT is analogous to classical absorption heat transformers…

Quantum Physics · Physics 2026-01-06 Arghya Maity , Paranjoy Chaki , Ahana Ghoshal , Ujjwal Sen

This research employs the Kraus representation and Sz.-Nagy dilation theorem to model a three-level quantum heat on quantum circuits, investigating its dynamic evolution and thermodynamic performance. The feasibility of the dynamic model is…

Quantum Physics · Physics 2024-05-29 Gao-xiang Deng , Zhe He , Yu Liu , Wei Shao , Zheng Cui

Reinforcement learning (RL) systems have countless applications, from energy-grid management to protein design. However, such real-world scenarios are often extremely difficult, combinatorial in nature, and require complex coordination…

Wireless Body Area Networks (WBANs) have gained significant attention due to their applications in healthcare monitoring, sports, military communication, and remote patient care. These networks consist of wearable or implanted sensors that…

Cryptography and Security · Computer Science 2025-10-23 Abdollah Rahimi , Mehdi Jafari Shahbazzadeh , Amid Khatibi

A heat engine is a machine which uses the temperature difference between a hot and a cold reservoir to extract work. Here both reservoirs are quantum systems and a heat engine is described by a unitary transformation which decreases the…

Quantum Physics · Physics 2009-11-11 Dominik Janzing

This paper introduces the SOAR framework for imitation learning. SOAR is an algorithmic template that learns a policy from expert demonstrations with a primal dual style algorithm that alternates cost and policy updates. Within the policy…

Machine Learning · Computer Science 2025-06-02 Stefano Viel , Luca Viano , Volkan Cevher

We propose a scheme of multilayer thermoelectric engine where {\em one} electric current is coupled to {\em two} temperature gradients in three-terminal geometry. This is realized by resonant tunneling through quantum dots embedded in two…

Mesoscale and Nanoscale Physics · Physics 2015-07-06 Jian-Hua Jiang

We present PAT, a transformer-based network that learns complex temporal co-occurrence action dependencies in a video by exploiting multi-scale temporal features. In existing methods, the self-attention mechanism in transformers loses the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-10 Faegheh Sardari , Armin Mustafa , Philip J. B. Jackson , Adrian Hilton

In a partially observable Markov decision process (POMDP), an agent typically uses a representation of the past to approximate the underlying MDP. We propose to utilize a frozen Pretrained Language Transformer (PLT) for history…

Hydrogen's role is growing as an energy carrier, increasing the need for efficient production, with methane steam reforming being the most widely used technique. This process is crucial for applications like fuel cells, where hydrogen is…

Computational Engineering, Finance, and Science · Computer Science 2025-09-16 Zofia Pizoń , Shinji Kimijima , Grzegorz Brus

Modern power systems are experiencing a variety of challenges driven by renewable energy, which calls for developing novel dispatch methods such as reinforcement learning (RL). Evaluation of these methods as well as the RL agents are…

Systems and Control · Electrical Eng. & Systems 2022-09-22 Zekuan Yu , Guangchun Ruan , Xinyue Wang , Guanglun Zhang , Yiliu He , Haiwang Zhong

Most previous studies on multi-agent reinforcement learning focus on deriving decentralized and cooperative policies to maximize a common reward and rarely consider the transferability of trained policies to new tasks. This prevents such…

Machine Learning · Computer Science 2019-11-28 Heechang Ryu , Hayong Shin , Jinkyoo Park

Constrained Reinforcement Learning (RL) has emerged as a significant research area within RL, where integrating constraints with rewards is crucial for enhancing safety and performance across diverse control tasks. In the context of heating…

Machine Learning · Computer Science 2024-10-01 Baohe Zhang , Lilli Frison , Thomas Brox , Joschka Bödecker