English
Related papers

Related papers: Reinforcement Learning Based Goodput Maximization …

200 papers

We study the discrete-time linear-quadratic (LQ) control model using reinforcement learning (RL). Using entropy to measure the cost of exploration, we prove that the optimal feedback policy for the problem must be Gaussian type. Then, we…

Machine Learning · Statistics 2025-02-05 Lucky Li

We present a method that addresses the pain point of long lead-time required to deploy cell-level parameter optimisation policies to new wireless network sites. Given a sequence of action spaces represented by overlapping subsets of…

Machine Learning · Computer Science 2024-05-01 Cengis Hasan , Alexandros Agapitos , David Lynch , Alberto Castagna , Giorgio Cruciata , Hao Wang , Aleksandar Milenovic

The functionality of Large Language Model (LLM) agents is primarily determined by two capabilities: action planning and answer summarization. The former, action planning, is the core capability that dictates an agent's performance. However,…

Machine Learning · Computer Science 2025-08-28 Zhiwei Li , Yong Hu , Wenqing Wang

Distributed optimal control is known to be challenging and can become intractable even for linear-quadratic regulator problems. In this work, we study a special class of such problems where distributed state feedback controllers can give…

Systems and Control · Electrical Eng. & Systems 2024-03-14 Johan Olsson , Runyu Zhang , Emma Tegling , Na Li

The optimization of electrical circuits is a difficult and time-consuming process performed by experts, but also increasingly by sophisticated algorithms. In this paper, a reinforcement learning (RL) approach is adapted to optimize a LLC…

Machine Learning · Computer Science 2023-03-02 Georg Kruse , Dominik Happel , Stefan Ditze , Stefan Ehrlich , Andreas Rosskopf

This article provides a methodology and open-source implementation of Reinforcement Learning algorithms for finding optimal routes in a packet-optical network scenario. The algorithm uses measurements provided by the physical layer (pre-FEC…

Networking and Internet Architecture · Computer Science 2024-06-24 A. L. García Navarro , Nataliia Koneva , Alfonso Sánchez-Macián , José Alberto Hernández , Óscar González de Dios , J. M. Rivas-Moscoso

Ultra-reliable and low-latency communications (URLLC) is firstly proposed in 5G networks, and expected to support applications with the most stringent quality-of-service (QoS). However, since the wireless channels vary dynamically, the…

Signal Processing · Electrical Eng. & Systems 2023-05-26 Kang Li , Pengcheng Zhu , Yan Wang , Fu-Chun Zheng , Xiaohu You

Quantum reservoir computing (QRC) leverages the natural dynamics of quantum systems to process time-series data efficiently, offering a promising approach for near-term quantum devices. Unlike classical reservoir computing, the efficacy of…

Quantum Physics · Physics 2025-03-25 Tomoya Monomi , Wataru Setoyama , Yoshihiko Hasegawa

This paper investigates the effective capacity of a point-to-point ultra-reliable low latency communication (URLLC) transmission over multiple parallel sub-channels at finite blocklength (FBL) with imperfect channel state information (CSI).…

Information Theory · Computer Science 2022-10-20 Hongsen Peng , Meixia Tao

Large language models (LLMs) are shifting from answer providers to intelligent tutors in educational settings, yet current supervised fine-tuning methods only learn surface teaching patterns without dynamic adaptation capabilities. Recent…

Artificial Intelligence · Computer Science 2026-01-06 Shouang Wei , Min Zhang , Xin Lin , Bo Jiang , Kun Kuang , Zhongxiang Dai

A supervised learning approach for the solution of large-scale nonlinear stabilization problems is presented. A stabilizing feedback law is trained from a dataset generated from State-dependent Riccati Equation solves. The training phase is…

Optimization and Control · Mathematics 2021-03-09 Giacomo Albi , Sara Bicego , Dante Kalise

Measurement is an essential component of robust and practical quantum computation. For superconducting qubits, the measurement process involves the effective manipulation of the joint qubit-resonator dynamics, and it should ideally provide…

Quantum Physics · Physics 2025-07-10 Aniket Chatterjee , Jonathan Schwinger , Yvonne Y. Gao

Ultra-reliable and low-latency communication (URLLC) is a pivotal technique for enabling the wireless control over industrial Internet-of-Things (IIoT) devices. By deploying distributed access points (APs), cell-free massive multiple-input…

Information Theory · Computer Science 2023-02-21 Qihao Peng , Hong Ren , Cunhua Pan , Nan Liu , Maged Elkashlan

Ultra-reliable and low-latency communications (URLLC) are considered as one of three new application scenarios in the fifth generation cellular networks. In this work, we aim to reduce the user experienced delay through prediction and…

Signal Processing · Electrical Eng. & Systems 2024-10-30 Zhanwei Hou , Changyang She , Yonghui Li , Zhuo Li , Branka Vucetic

Frequency control is an important problem in modern recommender systems. It dictates the delivery frequency of recommendations to maintain product quality and efficiency. For example, the frequency of delivering promotional notifications…

Machine Learning · Computer Science 2020-12-22 Yang Liu , Zhengxing Chen , Kittipat Virochsiri , Juan Wang , Jiahao Wu , Feng Liang

The importance of feedback control is being increasingly appreciated in quantum physics and applications. This paper describes the use of optimal control methods in the design of quantum feedback control systems, and in particular the paper…

Quantum Physics · Physics 2009-11-10 M. R. James

Meta-reinforcement learning (meta-RL) acquires meta-policies that show good performance for tasks in a wide task distribution. However, conventional meta-RL, which learns meta-policies by randomly sampling tasks, has been reported to show…

Machine Learning · Computer Science 2022-04-01 Morio Matsumoto , Hiroya Matsuba , Toshihiro Kujirai

Reinforcement learning (RL) enables agents to learn optimal policies through environmental interaction. However, RL suffers from reduced learning efficiency due to the curse of dimensionality in high-dimensional spaces. Quantum…

Machine Learning · Computer Science 2025-07-02 Seok Bin Son , Joongheon Kim

Quantum error correction is widely thought to be the key to fault-tolerant quantum computation. However, determining the most suited encoding for unknown error channels or specific laboratory setups is highly challenging. Here, we present a…

We investigate the role of information in active feedback control of quantum many-body systems using reinforcement learning. Active feedback breaks detailed balance, enabling the engineering of steady states and dynamical phases of matter…

Quantum Physics · Physics 2025-08-12 Giovanni Cemin , Markus Schmitt , Marin Bukov
‹ Prev 1 4 5 6 7 8 10 Next ›