中文
相关论文

相关论文: A Reinforcement Learning based approach for Multi-…

200 篇论文

To avoid myopic behavior, multi-step lookahead Bayesian optimization (BO) algorithms consider the sequential nature of BO and have demonstrated promising results in recent years. However, owing to the curse of dimensionality, most of these…

机器学习 · 计算机科学 2026-04-24 Mujin Cheon , Jay H. Lee , Dong-Yeun Koh , Calvin Tsay

Effective traffic control is essential for mitigating congestion in transportation networks. Conventional traffic management strategies, including route guidance and ramp metering, often rely on state feedback controllers, which are used…

机器学习 · 计算机科学 2026-04-13 Giray Önür , Azita Dabiri , Bart De Schutter

Fine-tuning foundation models has emerged as a powerful approach for generating objects with specific desired properties. Reinforcement learning (RL) provides an effective framework for this purpose, enabling models to generate outputs that…

机器学习 · 计算机科学 2025-11-04 Pouya M. Ghari , Simone Sciabola , Ye Wang

The paper presents a joint beamforming algorithm using statistical channel state information (S-CSI) for reconfigurable intelligent surfaces (RIS) for multiuser MISO wireless communications. We used S-CSI, which is a long-term average of…

信号处理 · 电气工程与系统科学 2022-09-21 Mahdi Eskandari , Huiling Zhu , Arman Shojaeifard , Jiangzhou Wang

Designing a cognitive radar system capable of adapting its parameters is challenging, particularly when tasked with tracking a ballistic missile throughout its entire flight. In this work, we focus on proposing adaptive algorithms that…

信号处理 · 电气工程与系统科学 2024-10-15 Thulasi Tholeti , Avinash Rangarajan , Sheetal Kalyani

We consider the design of polyphase waveforms for ground moving target detection with airborne multiple-input-multiple-output (MIMO) radar. Due to the constant-modulus and finite-alphabet constraint on the waveforms, the associated design…

信号处理 · 电气工程与系统科学 2020-10-28 Bo Tang , Jonathan Tuck , Peter Stoica

Reinforcement learning (RL) is a promising tool to solve robust optimal well control problems where the model parameters are highly uncertain, and the system is partially observable in practice. However, RL of robust control policies often…

机器学习 · 计算机科学 2022-07-14 Atish Dixit , Ahmed H. ElSheikh

Underwater target localization using range-only and single-beacon (ROSB) techniques with autonomous vehicles has been used recently to improve the limitations of more complex methods, such as long baseline and ultra-short baseline systems.…

机器人学 · 计算机科学 2023-01-18 Ivan Masmitja , Mario Martin , Kakani Katija , Spartacus Gomariz , Joan Navarro

This paper addresses the design issues of the multi-antenna-based cognitive radio (CR) system that is able to operate concurrently with the licensed primary radio (PR) system. We propose a practical CR transmission strategy consisting of…

信息论 · 计算机科学 2009-05-12 Feifei Gao , Rui Zhang , Ying-Chang Liang , Xiaodong Wang

The combination of multiple-input multiple-output (MIMO) systems and intelligent reflecting surfaces (IRSs) is foreseen as a critical enabler of beyond 5G (B5G) and 6G. In this work, two different approaches are considered for the joint…

信息论 · 计算机科学 2024-01-31 Dariel Pereira-Ruisánchez , Óscar Fresnedo , Darian Pérez-Adán , Luis Castedo

The increasing scale of manycore systems poses significant challenges in managing reliability while meeting performance demands. Simultaneously, these systems become more susceptible to different aging mechanisms such as negative-bias…

机器学习 · 计算机科学 2024-12-30 Fatemeh Hossein-Khani , Omid Akbari

The electromagnetic inverse problem has long been a research hotspot. This study aims to reverse radar view angles in synthetic aperture radar (SAR) images given a target model. Nonetheless, the scarcity of SAR data, combined with the…

机器学习 · 计算机科学 2024-01-03 Yanni Wang , Hecheng Jia , Shilei Fu , Huiping Lin , Feng Xu

Mobile users are prone to experience beam failure due to beam drifting in millimeter wave (mmWave) communications. Sensing can help alleviate beam drifting with timely beam changes and low overhead since it does not need user feedback. This…

信号处理 · 电气工程与系统科学 2025-05-12 Xiyu Wang , Gilberto Berardinelli , Hei Victor Cheng , Petar Popovski , Ramoni Adeogun

Reinforcement learning (RL) solves sequential decision-making problems via a trial-and-error process interacting with the environment. While RL achieves outstanding success in playing complex video games that allow huge trial-and-error,…

机器学习 · 计算机科学 2022-06-22 Fan-Ming Luo , Tian Xu , Hang Lai , Xiong-Hui Chen , Weinan Zhang , Yang Yu

Model-based reinforcement learning (RL) is a sample-efficient way of learning complex behaviors by leveraging a learned single-step dynamics model to plan actions in imagination. However, planning every action for long-horizon tasks is not…

机器学习 · 计算机科学 2022-12-13 Lucy Xiaoyang Shi , Joseph J. Lim , Youngwoon Lee

Reinforcement learning (RL) algorithms aim to learn optimal decisions in unknown environments through experience of taking actions and observing the rewards gained. In some cases, the environment is not influenced by the actions of the RL…

This paper develops an active sensing framework for designing the transmit and receive beamformers of a multiple-input multiple-output (MIMO) radar system. In the proposed technique, the beamformers are adaptively designed in each sensing…

信号处理 · 电气工程与系统科学 2026-04-21 Nadim Ghaddar , Kareem M. Attiah , Wei Yu

This study focuses on a multi-user massive multiple-input multiple-output (MU-mMIMO) system by incorporating an unmanned aerial vehicle (UAV) as a decode-and-forward (DF) relay between the base station (BS) and multiple Internet-of-Things…

信号处理 · 电气工程与系统科学 2024-04-11 MohammadMahdi Ghadaksaz , Mobeen Mahmood , Tho Le-Ngoc

Reinforcement Learning (RL) is a potent tool for sequential decision-making and has achieved performance surpassing human capabilities across many challenging real-world tasks. As the extension of RL in the multi-agent system domain,…

In the field of autonomous Unmanned Aerial Vehicles (UAVs) landing, conventional approaches fall short in delivering not only the required precision but also the resilience against environmental disturbances. Yet, learning-based algorithms…

计算机视觉与模式识别 · 计算机科学 2024-05-22 Francisco Neves , Luís Branco , Maria Pereira , Rafael Claro , Andry Pinto