中文
相关论文

相关论文: Bayesian Nonparametric Reinforcement Learning in L…

200 篇论文

The main challenge of multiagent reinforcement learning is the difficulty of learning useful policies in the presence of other simultaneously learning agents whose changing behaviors jointly affect the environment's transition and reward…

Federated Bayesian learning offers a principled framework for the definition of collaborative training algorithms that are able to quantify epistemic uncertainty and to produce trustworthy decisions. Upon the completion of collaborative…

机器学习 · 计算机科学 2021-04-09 Jinu Gong , Osvaldo Simeone , Joonhyuk Kang

We consider the issue of multiple agents learning to communicate through reinforcement learning within partially observable environments, with a focus on information asymmetry in the second part of our work. We provide a review of the…

机器学习 · 计算机科学 2019-11-14 Mohamed Salah Zaïem , Etienne Bennequin

The real world is unpredictable. Therefore, to solve long-horizon decision-making problems with autonomous robots, we must construct agents that are capable of adapting to changes in the environment during deployment. Model-based planning…

机器人学 · 计算机科学 2024-10-01 Alicia Li , Nishanth Kumar , Tomás Lozano-Pérez , Leslie Kaelbling

We address the problem of Bayesian reinforcement learning using efficient model-based online planning. We propose an optimism-free Bayes-adaptive algorithm to induce deeper and sparser exploration with a theoretical bound on its performance…

机器学习 · 计算机科学 2020-06-30 Divya Grover , Debabrota Basu , Christos Dimitrakakis

We develop a framework for spectrum sensing in cooperative amplify-and-forward cognitive radio networks. We consider a stochastic model where relays are assigned in cognitive radio networks to transmit the primary user's signal to a…

信息论 · 计算机科学 2011-11-02 Ido Nevat , Gareth W. Peters , Jinhong Yuan , Iain Collings

Multiagent systems provide an ideal environment for the evaluation and analysis of real-world problems using reinforcement learning algorithms. Most traditional approaches to multiagent learning are affected by long training periods as well…

人工智能 · 计算机科学 2021-05-25 Unnikrishnan Rajendran Menon , Anirudh Rajiv Menon

The unlicensed spectrum has been utilized to make up the shortage on frequency spectrum in new radio (NR) systems. To fully exploit the advantages brought by the unlicensed bands, one of the key issues is to guarantee the fair coexistence…

信息论 · 计算机科学 2021-02-24 Rui Yin , Zhiqun Zou , Celimuge Wu , Jiantao Yuan , Xianfu Chen , Guanding Yu

This paper proposes a safe reinforcement learning algorithm for generation bidding decisions and unit maintenance scheduling in a competitive electricity market environment. In this problem, each unit aims to find a bidding strategy that…

系统与控制 · 电气工程与系统科学 2021-12-21 Pegah Rokhforoz , Olga Fink

A network of agents attempt to learn some unknown state of the world drawn by nature from a finite set. Agents observe private signals conditioned on the true state, and form beliefs about the unknown state accordingly. Each agent may face…

机器学习 · 计算机科学 2015-03-13 Shahin Shahrampour , Mohammad Amin Rahimian , Ali Jadbabaie

In this paper, we study how to solve resource allocation problems in ultra-reliable and low-latency communications by unsupervised deep learning, which often yield functional optimization problems with quality-of-service (QoS) constraints.…

网络与互联网体系结构 · 计算机科学 2019-06-06 Chengjian Sun , Chenyang Yang

Mobile networks are composed of many base stations and for each of them many parameters must be optimized to provide good services. Automatically and dynamically optimizing all these entities is challenging as they are sensitive to…

机器学习 · 计算机科学 2021-10-01 Maxime Bouton , Hasan Farooq , Julien Forgeat , Shruti Bothe , Meral Shirazipour , Per Karlsson

Data-efficient learning algorithms are essential in many practical applications for which data collection is expensive, e.g., for the optimal deployment of wireless systems in unknown propagation scenarios. Meta-learning can address this…

机器学习 · 计算机科学 2022-05-25 Ivana Nikoloska , Osvaldo Simeone

With the exponential increase in mobile users, the mobile data demand has grown tremendously. To meet these demands, cellular operators are constantly innovating to enhance the capacity of cellular systems. Consequently, operators have been…

网络与互联网体系结构 · 计算机科学 2021-12-30 Vanlin Sathya , Srikant Manas Kala , Kalpana Naidu

Based on the License-Assisted Access (LAA) small cell architecture, the LAA coexisting with Wi-Fi heterogeneous networks provides LTE mobile users with high bandwidth efficiency as the unlicensed channels are shared among LAA and Wi-Fi.…

网络与互联网体系结构 · 计算机科学 2025-09-30 Po-Heng Chou

In this paper, a proactive dynamic spectrum sharing scheme between 4G and 5G systems is proposed. In particular, a controller decides on the resource split between NR and LTE every subframe while accounting for future network states such as…

网络与互联网体系结构 · 计算机科学 2021-02-23 Ursula Challita , David Sandberg

Reinforcement learning for LLM agents is typically conducted on a static data distribution, which fails to adapt to the agent's evolving behavior and leads to poor coverage of complex environment interactions. To address these challenges,…

计算与语言 · 计算机科学 2026-04-20 Shidong Yang , Ziyu Ma , Tongwen Huang , Yiming Hu , Yong Wang , Xiangxiang Chu

With the rise of online e-commerce platforms, more and more customers prefer to shop online. To sell more products, online platforms introduce various modules to recommend items with different properties such as huge discounts. A web page…

机器学习 · 计算机科学 2020-09-01 Xu He , Bo An , Yanghua Li , Haikai Chen , Rundong Wang , Xinrun Wang , Runsheng Yu , Xin Li , Zhirong Wang

LTE-Licensed Assisted Access (LAA) networks are beginning to be deployed widely in major metropolitan areas in the US in the unlicensed 5 GHz bands, which have existing dense deployments of Wi-Fi. This provides a real-world opportunity to…

网络与互联网体系结构 · 计算机科学 2021-04-02 Vanlin Sathya , Muhammad Iqbal Rochman , Monisha Ghosh

This paper presents a novel deep reinforcement learning-based resource allocation technique for the multi-agent environment presented by a cognitive radio network that coexists through underlay dynamic spectrum access (DSA) with a primary…

网络与互联网体系结构 · 计算机科学 2020-03-09 Ankita Tondwalkar , Dr Andres Kwasinski