中文
相关论文

相关论文: Multi-Agent Continuous Control with Generative Flo…

200 篇论文

In our prior work we have proposed the use of GFlowNets, a generative AI (GenAI) framework, for designing a secure communication system comprising a time-modulated intelligent reflecting surface (TM-IRS). However, GFlowNet-based approaches…

信号处理 · 电气工程与系统科学 2025-11-11 Zhihao Tao , Athina P. Petropulu

Real-time traffic flow prediction holds significant importance within the domain of Intelligent Transportation Systems (ITS). The task of achieving a balance between prediction precision and computational efficiency presents a significant…

机器学习 · 计算机科学 2024-04-08 Muhammad Yaqub , Shahzad Ahmad , Malik Abdul Manan , Imran Shabir Chuhan

Reinforcement learning has gained traction for active flow control tasks, with initial applications exploring drag mitigation via flow field augmentation around a two-dimensional cylinder. RL has since been extended to more complex…

机器学习 · 计算机科学 2025-04-01 Marius Kurz , Rohan Kaushik , Marcel Blind , Patrick Kopper , Anna Schwarz , Felix Rodach , Andrea Beck

Multiagent systems consist of agents that locally exchange information through a physical network subject to a graph topology. Current control methods for networked multiagent systems assume the knowledge of graph topologies in order to…

最优化与控制 · 数学 2015-01-15 Tansel Yucelen , John Daniel Peterson , Kevin L. Moore

Accurate and refined passenger flow prediction is essential for optimizing the collaborative management of multiple collection and distribution modes in large-scale transportation hubs. Traditional methods often focus only on the overall…

机器学习 · 计算机科学 2025-04-10 Ronghui Zhang , Wenbin Xing , Mengran Li , Zihan Wang , Junzhou Chen , Xiaolei Ma , Zhiyuan Liu , Zhengbing He

In this paper, we explore the use of multi-agent deep learning as well as learning to cooperate principles to meet stringent service level agreements, in terms of throughput and end-to-end delay, for a set of classified network flows. We…

网络与互联网体系结构 · 计算机科学 2022-05-25 Hassan Fawaz , Julien Lesca , Pham Tran Anh Quang , Jérémie Leguay , Djamal Zeghlache , Paolo Medagliani

In multi-agent reinforcement learning, a commonly considered paradigm is centralized training with decentralized execution. However, in this framework, decentralized execution restricts the development of coordinated policies due to the…

多智能体系统 · 计算机科学 2024-12-30 Wenzhe Fan , Zishun Yu , Chengdong Ma , Changye Li , Yaodong Yang , Xinhua Zhang

Controlling generative models is computationally expensive. This is because optimal alignment with a reward function--whether via inference-time steering or fine-tuning--requires estimating the value function. This task demands access to…

Multi-agent reinforcement learning faces fundamental challenges that conventional approaches have failed to overcome: exponentially growing joint action spaces, non-stationary environments where simultaneous learning creates moving targets,…

人工智能 · 计算机科学 2025-07-15 Hang Wang , Junshan Zhang

Training deep neural networks remains computationally intensive due to the itera2 tive nature of gradient-based optimization. We propose Gradient Flow Matching (GFM), a continuous-time modeling framework that treats neural network training…

机器学习 · 计算机科学 2025-05-27 Xiao Shou , Yanna Ding , Jianxi Gao

We tackle the problem of sampling from intractable high-dimensional density functions, a fundamental task that often appears in machine learning and statistics. We extend recent sampling-based approaches that leverage controlled stochastic…

机器学习 · 计算机科学 2024-03-12 Dinghuai Zhang , Ricky T. Q. Chen , Cheng-Hao Liu , Aaron Courville , Yoshua Bengio

Generative models in drug discovery have recently gained attention as efficient alternatives to brute-force virtual screening. However, most existing models do not account for synthesizability, limiting their practical use in real-world…

生物大分子 · 定量生物学 2025-03-07 Seonghwan Seo , Minsu Kim , Tony Shen , Martin Ester , Jinkyoo Park , Sungsoo Ahn , Woo Youn Kim

Multi-Agent Path Finding (MAPF) seeks collision-free paths for multiple agents from their respective starting locations to their respective goal locations while minimizing path costs. Although many MAPF algorithms were developed and can…

多智能体系统 · 计算机科学 2024-12-24 Shuai Zhou , Shizhe Zhao , Zhongqiang Ren

In computational reinforcement learning, a growing body of work seeks to construct an agent's perception of the world through predictions of future sensations; predictions about environment observations are used as additional input features…

机器学习 · 计算机科学 2022-06-15 Alexandra Kearney , Anna Koop , Johannes Günther , Patrick M. Pilarski

Real-world fine-tuning of dexterous manipulation policies remains challenging due to limited real-world interaction budgets and highly multimodal action distributions. Diffusion-based policies, while expressive, do not permit conservative…

机器人学 · 计算机科学 2026-04-07 Chenyu Yang , Denis Tarasov , Davide Liconti , Hehui Zheng , Robert K. Katzschmann

This paper builds bridges between two families of probabilistic algorithms: (hierarchical) variational inference (VI), which is typically used to model distributions over continuous spaces, and generative flow networks (GFlowNets), which…

机器学习 · 计算机科学 2023-03-03 Nikolay Malkin , Salem Lahlou , Tristan Deleu , Xu Ji , Edward Hu , Katie Everett , Dinghuai Zhang , Yoshua Bengio

Agentic recommendations cast recommenders as large language model (LLM) agents that can plan, reason, use tools, and interact with users of varying preferences in web applications. However, most existing agentic recommender systems focus on…

计算与语言 · 计算机科学 2026-01-27 Yu Xia , Sungchul Kim , Tong Yu , Ryan A. Rossi , Julian McAuley

In cooperative multi-agent reinforcement learning (MARL), combining value decomposition with actor-critic enables agents to learn stochastic policies, which are more suitable for the partially observable environment. Given the goal of…

机器学习 · 计算机科学 2023-02-13 Jiangxing Wang , Deheng Ye , Zongqing Lu

Graph Convolutional Networks (GCNs) have gained significant developments in representation learning on graphs. However, current GCNs suffer from two common challenges: 1) GCNs are only effective with shallow structures; stacking multiple…

机器学习 · 计算机科学 2019-12-13 Menghan Wang , Kun Zhang , Gulin Li , Keping Yang , Luo Si

Coordinating large populations of interacting agents is a central challenge in multi-agent reinforcement learning (MARL), where the size of the joint state-action space scales exponentially with the number of agents. Mean-field methods…

机器学习 · 计算机科学 2026-02-19 Emile Anand , Richard Hoffmann , Sarah Liaw , Adam Wierman
‹ 上一页 1 8 9 10 下一页 ›