English
Related papers

Related papers: Trend-Based SAC Beam Control Method with Zero-Shot…

200 papers

Fixed-frequency control in robotics imposes a trade-off between the efficiency of low-frequency control and the robustness of high-frequency control, a limitation not seen in adaptable biological systems. We address this with a…

Robotics · Computer Science 2025-10-28 Arnav Sukhija , Lenart Treven , Jin Cheng , Florian Dörfler , Stelian Coros , Andreas Krause

This paper introduces a sensing management method for integrated sensing and communications (ISAC) in cell-free massive multiple-input multiple-output (MIMO) systems. Conventional communication systems employ channel estimation procedures…

Signal Processing · Electrical Eng. & Systems 2025-10-09 Eren Berk Kama , Murat Babek Salman , Isaac Skog , Emil Björnson

This paper introduces a novel reinforcement learning (RL) strategy designed to facilitate rapid autonomy transfer by utilizing pre-trained critic value functions from multiple environments. Unlike traditional methods that require extensive…

We introduce a sequence-conditioned critic for Soft Actor--Critic (SAC) that models trajectory context with a lightweight Transformer and trains on aggregated $N$-step targets. Unlike prior approaches that (i) score state--action pairs in…

Machine Learning · Computer Science 2025-09-30 Dong Tian , Onur Celik , Gerhard Neumann

Triangular tethered formation system (TTFS) provide a promising platform for deep space exploration and distributed sensing due to its intrinsic spatial-orientation stability and capability of adjusting distances among node satellites…

Systems and Control · Electrical Eng. & Systems 2026-01-09 Xinyi Tao , Panfeng Huang , Fan Zhang

The Soft Actor-Critic (SAC) algorithm, a state-of-the-art method in maximum entropy reinforcement learning, traditionally relies on minimizing reverse Kullback-Leibler (KL) divergence for policy updates. However, this approach leads to an…

Machine Learning · Computer Science 2025-06-03 Yixian Zhang , Huaze Tang , Changxu Wei , Wenbo Ding

Safe reinforcement learning (RL) for robotic systems requires policies that improve task performance while satisfying state and input constraints during both training and deployment. Control barrier functions (CBFs) provide a principled…

Robotics · Computer Science 2026-05-27 Dhruv S. Kushwaha , Zoleikha A. Biron

Reinforcement Learning (RL) has been widely applied to many control tasks and substantially improved the performances compared to conventional control methods in many domains where the reward function is well defined. However, for many…

Machine Learning · Computer Science 2024-03-22 Baohe Zhang , Yuan Zhang , Lilli Frison , Thomas Brox , Joschka Bödecker

A high peak current, flat longitudinal phase space electron beam is desirable for efficient x-ray free electron laser (FEL) radiation in next generation light sources. To attain such a beam requires the extensive design of the linear…

Accelerator Physics · Physics 2019-10-23 Ji Qiang

This letter investigates the robust beamforming design for a near-field secure integrated sensing and communication (ISAC) system with multiple communication users (CUs) and targets, as well as multiple eavesdroppers. Taking into account…

Information Theory · Computer Science 2025-07-18 Ziqiang CHen , Feng Wang , Guojun Han , Xin Wang , Vincent K. N. Lau

State-of-the-art deep reinforcement learning (RL) methods have achieved remarkable performance in continuous control tasks, yet their computational complexity is often incompatible with the constraints of resource-limited hardware, due to…

Machine Learning · Computer Science 2026-05-12 Riccardo De Monte , Matteo Cederle , Gian Antonio Susto

Coordinated controlling a large UAV swarm requires significant spectrum resources due to the need for bandwidth allocation per UAV, posing a challenge in resource-limited environments. Over-the-air (OTA) control has emerged as a…

Signal Processing · Electrical Eng. & Systems 2025-02-12 Zhuangkun Wei , Wenxiu Hu , Yathreb Bouazizi , Mengbang Zou , Chenguang Liu , Yunfei Chen , Hongjian Sun , Julie McCann

Soft continuum arms (SCAs) soft and deformable nature presents challenges in modeling and control due to their infinite degrees of freedom and non-linear behavior. This work introduces a reinforcement learning (RL)-based framework for…

Robotics · Computer Science 2026-03-13 Hsin-Jung Yang , Mahsa Khosravi , Benjamin Walt , Girish Krishnan , Soumik Sarkar

In superconducting linear accelerators (linacs), accurately monitoring beam dynamics is essential for minimizing beam losses and ensuring stable operations. However, destructive diagnostics must be avoided in superconducting sections to…

Soft Actor-Critic (SAC) is one of the state-of-the-art off-policy reinforcement learning (RL) algorithms that is within the maximum entropy based RL framework. SAC is demonstrated to perform very well in a list of continous control tasks…

Machine Learning · Computer Science 2021-12-22 Zhenyang Shi , Surya P. N. Singh

This paper presents a centralized predictive cost adaptive control (PCAC) strategy for the position and attitude control of quadrotors. PCAC is an optimal, prediction-based control method that uses recursive least squares (RLS) to identify…

Systems and Control · Electrical Eng. & Systems 2025-08-26 Tam W. Nguyen

We present a Multi-task Soft Actor-Critic (SAC) Reinforcement Learning framework designed for open-system quantum control across diverse Hamiltonians, which learns optimal pulse sequences while simultaneously discovering problem-specific…

Quantum Physics · Physics 2026-05-27 Haftu W. Fentaw , Steve Campbell , Simon Caton

Integrated sensing and communication (ISAC) is a promising solution for the future sixth-generation (6G) system. However, classical fixed-position antenna (FPA) ISAC systems fail to fully utilize spatial degrees of freedom (DoFs), resulting…

Signal Processing · Electrical Eng. & Systems 2025-11-18 Yuan Guo , Wen Chen , Qingqing Wu , Yang Liu , Qiong Wu

In future 6G Mobile Edge Computing (MEC), autopilot systems require the capability of processing multimodal data with strong interdependencies. However, traditional heuristic algorithms are inadequate for real-time scheduling due to their…

Networking and Internet Architecture · Computer Science 2023-07-17 Ke Deng , Zhiyuan He , Hao Zhang , Haohan Lin , Desheng Wang

Variable stiffness actuator (VSA) designs are manifold. Conventional model-based control of these nonlinear systems is associated with high effort and design-dependent assumptions. In contrast, machine learning offers a promising…

Systems and Control · Electrical Eng. & Systems 2024-04-17 Tim-Lukas Habich , Sarah Kleinjohann , Moritz Schappler