English
Related papers

Related papers: RCBSF: A Multi-Agent Framework for Automated Contr…

200 papers

We explore deep Reinforcement Learning(RL) algorithms for scalping trading and knew that there is no appropriate trading gym and agent examples. Thus we propose gym and agent like Open AI gym in finance. Not only that, we introduce new RL…

Artificial Intelligence · Computer Science 2019-04-02 Uk Jo , Taehyun Jo , Wanjun Kim , Iljoo Yoon , Dongseok Lee , Seungho Lee

Agent-based models (ABMs) are valuable for modelling complex, potentially out-of-equilibria scenarios. However, ABMs have long suffered from the Lucas critique, stating that agent behaviour should adapt to environmental changes.…

Multiagent Systems · Computer Science 2025-01-17 Benjamin Patrick Evans , Sihan Zeng , Sumitra Ganesh , Leo Ardon

Automated \enquote{LLM-as-a-Judge} frameworks have become the de facto standard for scalable evaluation across natural language processing. For instance, in safety evaluation, these judges are relied upon to evaluate harmfulness in order to…

Computation and Language · Computer Science 2026-03-17 Leo Schwinn , Moritz Ladenburger , Tim Beyer , Mehrnaz Mofakhami , Gauthier Gidel , Stephan Günnemann

Recent research efforts indicate that federated learning (FL) systems are vulnerable to a variety of security breaches. While numerous defense strategies have been suggested, they are mainly designed to counter specific attack patterns and…

Cryptography and Security · Computer Science 2025-12-19 Henger Li , Tianyi Xu , Tao Li , Yunian Pan , Quanyan Zhu , Zizhan Zheng

Multi-Agent Reinforcement Learning (MARL) algorithms show amazing performance in simulation in recent years, but placing MARL in real-world applications may suffer safety problems. MARL with centralized shields was proposed and verified in…

Multiagent Systems · Computer Science 2021-03-24 Zhiyuan Cai , Huanhui Cao , Wenjie Lu , Lin Zhang , Hao Xiong

Recent work, spanning from autonomous vehicle coordination to in-space assembly, has shown the importance of learning collaborative behavior for enabling robots to achieve shared goals. A common approach for learning this cooperative…

Multiagent Systems · Computer Science 2025-02-25 Kartik Nagpal , Dayi Dong , Jean-Baptiste Bouvier , Negar Mehr

Guided cooperation allows intelligent agents with heterogeneous capabilities to work together by following a leader-follower type of interaction. However, the associated control problem becomes challenging when the leader agent does not…

Systems and Control · Electrical Eng. & Systems 2024-02-01 Yuhan Zhao , Quanyan Zhu

Optimal control methods provide solutions to safety-critical problems but easily become intractable. Control Barrier Functions (CBFs) have emerged as a popular technique that facilitates their solution by provably guaranteeing safety,…

Systems and Control · Electrical Eng. & Systems 2025-02-21 Ehsan Sabouni , H. M. Sabbir Ahmad , Vittorio Giammarino , Christos G. Cassandras , Ioannis Ch. Paschalidis , Wenchao Li

Ensuring safety in MARL, particularly when deploying it in real-world applications such as autonomous driving, emerges as a critical challenge. To address this challenge, traditional safe MARL methods extend MARL approaches to incorporate…

Robotics · Computer Science 2024-05-29 Zhi Zheng , Shangding Gu

Motivated by the omnipresence of hierarchical structures in many real-world applications, this study delves into the intricate realm of bi-level games, with a specific focus on exploring local Stackelberg equilibria as a solution concept.…

Systems and Control · Electrical Eng. & Systems 2024-02-23 Marko Maljkovic , Gustav Nilsson , Nikolas Geroliminis

Automated code generation has long been considered the holy grail of software engineering. The emergence of Large Language Models (LLMs) has catalyzed a revolutionary breakthrough in this area. However, existing methods that only rely on…

Software Engineering · Computer Science 2025-08-27 Xu Lu , Weisong Sun , Yiran Zhang , Ming Hu , Cong Tian , Zhi Jin , Yang Liu

The automated repair of C++ compilation errors presents a significant challenge, the resolution of which is critical for developer productivity. Progress in this domain is constrained by two primary factors: the scarcity of large-scale,…

Artificial Intelligence · Computer Science 2025-09-22 Weixuan Sun , Jucai Zhai , Dengfeng Liu , Xin Zhang , Xiaojun Wu , Qiaobo Hao , AIMgroup , Yang Fang , Jiuyang Tang

Iterative code generation with Large Language Models (LLMs) can be viewed as an optimization process guided by textual feedback. However, existing LLM self-correction methods predominantly operate in a stateless, trial-and-error manner akin…

Machine Learning · Computer Science 2026-03-31 Zizheng Zhang , Yuyang Liao , Chen Chen , Jian He , Dun Wu , Qianjin Yu , Yanqin Gao , Jin Yang , Kailai Zhang , Eng Siong Chng , Xionghu Zhong

We present CRM (Multi-Agent Collaborative Reward Model), a framework that replaces a single black-box reward model with a coordinated team of specialist evaluators to improve robustness and interpretability in RLHF. Conventional reward…

Artificial Intelligence · Computer Science 2026-01-06 Pei Yang , Ke Zhang , Ji Wang , Xiao Chen , Yuxin Tang , Eric Yang , Lynn Ai , Bill Shi

Ensuring robot safety in complex environments is a difficult task due to actuation limits, such as torque bounds. This paper presents a safety-critical control framework that leverages learning-based switching between multiple backup…

Robotics · Computer Science 2024-03-08 Neil C. Janwani , Ersin Daş , Thomas Touma , Skylar X. Wei , Tamas G. Molnar , Joel W. Burdick

Control Barrier Functions (CBFs) have become powerful tools for ensuring safety in nonlinear systems. However, finding valid CBFs that guarantee persistent safety and feasibility remains an open challenge, especially in systems with input…

Robotics · Computer Science 2025-03-05 Taekyung Kim , Robin Inho Kee , Dimitra Panagou

This paper presents a decentralized safety filter for collision avoidance in multi-agent aerospace interception scenarios. The approach leverages robust control barrier functions (RCBFs) to guarantee forward invariance of safety sets under…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Johannes Autenrieb , Mark Spiller

Vertical federated learning (VFL) attracts increasing attention due to the emerging demands of multi-party collaborative modeling and concerns of privacy leakage. In the real VFL applications, usually only one or partial parties hold…

Machine Learning · Computer Science 2021-03-02 Qingsong Zhang , Bin Gu , Cheng Deng , Heng Huang

Artificial behavioral agents are often evaluated based on their consistent behaviors and performance to take sequential actions in an environment to maximize some notion of cumulative reward. However, human decision making in real life…

Artificial Intelligence · Computer Science 2021-12-28 Baihan Lin , Guillermo Cecchi , Djallel Bouneffouf , Jenna Reinen , Irina Rish

Real-Time Bidding (RTB) enables advertisers to place competitive bids on impression opportunities instantaneously, striving for cost-effectiveness in a highly competitive landscape. Although RTB has widely benefited from the utilization of…

Artificial Intelligence · Computer Science 2025-02-04 Leng Cai , Junxuan He , Yikai Li , Junjie Liang , Yuanping Lin , Ziming Quan , Yawen Zeng , Jin Xu