中文
相关论文

相关论文: Reinforcement Learning Based Goodput Maximization …

200 篇论文

We investigate remote estimation over a Gilbert-Elliot channel with feedback. We assume that the channel state is observed by the receiver and fed back to the transmitter with one unit delay. In addition, the transmitter gets ACK/NACK…

最优化与控制 · 数学 2017-01-24 Jhelum Chakravorty , Aditya Mahajan

Large language models (LLMs) have achieved remarkable success in a wide range of natural language processing tasks and can be adapted through prompting. However, they remain suboptimal in multi-turn interactions, often relying on incorrect…

Vehicular crowdsensing is anticipated to become a key catalyst for data-driven optimization in the Intelligent Transportation System (ITS) domain. Yet, the expected growth in massive Machine-type Communication (mMTC) caused by…

网络与互联网体系结构 · 计算机科学 2020-01-16 Benjamin Sliwa , Christian Wietfeld

The ability of reinforcement learning algorithms to learn effective policies is determined by the rewards available during training. However, for practical problems, obtaining large quantities of reward labels is often infeasible due to…

机器学习 · 计算机科学 2025-10-02 Shreyas Chaudhari , Renhao Zhang , Philip S. Thomas , Bruno Castro da Silva

Modelling the electrical response of multi-level quantum systems at finite frequency has been typically performed in the context of two incomplete paradigms: (i) input-output theory, which is valid at any frequency but neglects dynamic…

介观与纳米尺度物理 · 物理学 2025-10-14 L. Peri , M. Benito , C. J. B. Ford , M. F. Gonzalez-Zalba

Equalizer parameter optimization is critical for signal integrity in high-speed memory systems operating at multi-gigabit data rates. However, existing methods suffer from computationally expensive eye diagram evaluation, optimization of…

机器学习 · 计算机科学 2026-05-07 Muhammad Usama , Dong Eui Chang

We consider a MIMO fading broadcast channel where the fading channel coefficients are constant over time-frequency blocks that span a coherent time $\times$ a coherence bandwidth. In closed-loop systems, channel state information at…

信息论 · 计算机科学 2009-12-11 Mari Kobayashi , Nihar Jindal , Giuseppe Caire

As one of the key communication scenarios in the 5th and also the 6th generation (6G) of mobile communication networks, ultra-reliable and low-latency communications (URLLC) will be central for the development of various emerging…

信号处理 · 电气工程与系统科学 2021-01-21 Changyang She , Chengjian Sun , Zhouyou Gu , Yonghui Li , Chenyang Yang , H. Vincent Poor , Branka Vucetic

Visible light communication (VLC) has been widely applied as a promising solution for modern short range communication. When it comes to the deployment of LED arrays in VLC networks, the emerging ultra-dense network (UDN) technology can be…

信号处理 · 电气工程与系统科学 2023-03-10 Xiao Tang , Sicong Liu

Multicasting in wireless systems is a natural way to exploit the redundancy in user requests in a Content Centric Network. Power control and optimal scheduling can significantly improve the wireless multicast network's performance under…

网络与互联网体系结构 · 计算机科学 2021-12-08 Ramkumar Raghu , Mahadesh Panju , Vaneet Aggarwal , Vinod Sharma

Reinforcement learning for the optimization of quantum circuits uses an agent whose goal is to maximize the value of a reward function that decides what is correct and what is wrong during the exploration of the search space. It is an open…

量子物理 · 物理学 2023-11-22 Ioana Moflic , Alexandru Paler

This paper considers a class of reinforcement learning problems, which involve systems with two types of states: stochastic and pseudo-stochastic. In such systems, stochastic states follow a stochastic transition kernel while the…

机器学习 · 计算机科学 2023-11-09 Honghao Wei , Xin Liu , Weina Wang , Lei Ying

With the rapid deployment of the Internet of Things (IoT), fifth-generation (5G) and beyond 5G networks are required to support massive access of a huge number of devices over limited radio spectrum radio. In wireless networks, different…

信号处理 · 电气工程与系统科学 2020-12-18 Helin Yang , Zehui Xiong , Jun Zhao , Dusit Niyato , Chau Yuen , Ruilong Deng

Digital quantum simulation is a promising application for quantum computers. Their free programmability provides the potential to simulate the unitary evolution of any many-body Hamiltonian with bounded spectrum by discretizing the time…

量子物理 · 物理学 2021-09-15 Adrien Bolens , Markus Heyl

A major challenge in the field of education is providing review schedules that present learned items at appropriate intervals to each student so that memory is retained over time. In recent years, attempts have been made to formulate item…

人工智能 · 计算机科学 2021-08-03 Yoshiki Kubotani , Yoshihiro Fukuhara , Shigeo Morishima

Conversion rate prediction is critical to many online applications such as digital display advertising. To capture dynamic data distribution, industrial systems often require retraining models on recent data daily or weekly. However, the…

信息检索 · 计算机科学 2023-07-25 Yifan Wang , Peijie Sun , Min Zhang , Qinglin Jia , Jingjie Li , Shaoping Ma

We propose to use channel inversion power control (CIPC) to achieve one-way ultra-reliable and low-latency communications (URLLC), where only the transmission in one direction requires ultra reliability and low latency. Based on channel…

信息论 · 计算机科学 2022-02-22 Chunhui Li , Shihao Yan , Nan Yang , Xiangyun Zhou

This paper proposes a reinforcement learning-based approach for optimal transient frequency control in power systems with stability and safety guarantees. Building on Lyapunov stability theory and safety-critical control, we derive…

系统与控制 · 电气工程与系统科学 2024-02-22 Zhenyi Yuan , Changhong Zhao , Jorge Cortes

In this paper, the reinforcement learning (RL)-based optimal control problem is studied for multiplicative-noise systems, where input delay is involved and partial system dynamics is unknown. To solve a variant of Riccati-ZXL equations,…

最优化与控制 · 数学 2023-01-10 Hongxia Wang , Fuyu Zhao , Zhaorong Zhang , Juanjuan Xu , Xun Li

Reinforcement learning means learning a policy--a mapping of observations into actions--based on feedback from the environment. The learning can be viewed as browsing a set of policies while evaluating them by trial through interaction with…

机器学习 · 计算机科学 2017-05-25 Leonid Peshkin , Virginia Savova