中文
相关论文

相关论文: CyGIL: A Cyber Gym for Training Autonomous Agents …

200 篇论文

This work aims to enable autonomous agents for network cyber operations (CyOps) by applying reinforcement and deep reinforcement learning (RL/DRL). The required RL training environment is particularly challenging, as it must balance the…

人工智能 · 计算机科学 2023-04-05 Li Li , Jean-Pierre S. El Rami , Adrian Taylor , James Hailing Rao , Thomas Kunz

Autonomous cyber agents may be developed by applying reinforcement and deep reinforcement learning (RL/DRL), where agents are trained in a representative environment. The training environment must simulate with high-fidelity the network…

机器学习 · 计算机科学 2023-04-05 Li Li , Jean-Pierre S. El Rami , Adrian Taylor , James Hailing Rao , Thomas Kunz

Recently, reinforcement and deep reinforcement learning (RL/DRL) have been applied to develop autonomous agents for cyber network operations(CyOps), where the agents are trained in a representative environment using RL and particularly DRL…

密码学与安全 · 计算机科学 2023-09-12 Li Li , Jean-Pierre S. El Rami , Ryan Kerr , Adrian Taylor , Grant Vandenberghe

Simulated environments have proven invaluable in Autonomous Cyber Operations (ACO) where Reinforcement Learning (RL) agents can be trained without the computational overhead of emulation. These environments must accurately represent…

密码学与安全 · 计算机科学 2026-02-17 Konur Tholl , Mariam El Mezouar , Adrian Taylor , Ranwa Al Mallah

Reinforcement learning (RL) has been demonstrated suitable to develop agents that play complex games with human-level performance. However, it is not understood how to effectively use RL to perform cybersecurity tasks. To develop such…

密码学与安全 · 计算机科学 2021-03-16 Andres Molina-Markham , Cory Miniter , Becky Powell , Ahmad Ridley

Reinforcement Learning (RL) and Multi-Agent Reinforcement Learning (MARL) have emerged as promising methodologies for addressing challenges in automated cyber defence (ACD). These techniques offer adaptive decision-making capabilities in…

Technological trends show that Radio Frequency Reinforcement Learning (RFRL) will play a prominent role in the wireless communication systems of the future. Applications of RFRL range from military communications jamming to enhancing WiFi…

Autonomous Cyber Operations (ACO) involves the development of blue team (defender) and red team (attacker) decision-making agents in adversarial scenarios. To support the application of machine learning algorithms to solve this problem, and…

密码学与安全 · 计算机科学 2021-08-23 Maxwell Standen , Martin Lucas , David Bowman , Toby J. Richer , Junae Kim , Damian Marriott

In November 2025, the authors ran a workshop on the topic of what makes a good reinforcement learning (RL) environment for autonomous cyber defence (ACD). This paper details the knowledge shared by participants both during the workshop and…

Autonomous Cyber Operations (ACO) rely on Reinforcement Learning (RL) to train agents to make effective decisions in the cybersecurity domain. However, existing ACO applications require agents to learn from scratch, leading to slow…

机器学习 · 计算机科学 2025-08-21 Konur Tholl , Mariam El Mezouar , Ranwa Al Mallah

We introduce ComputerRL, a framework for autonomous desktop intelligence that enables agents to operate complex digital workspaces skillfully. ComputerRL features the API-GUI paradigm, which unifies programmatic API calls and direct GUI…

人工智能 · 计算机科学 2025-10-22 Hanyu Lai , Xiao Liu , Yanxiao Zhao , Han Xu , Hanchen Zhang , Bohao Jing , Yanyu Ren , Shuntian Yao , Yuxiao Dong , Jie Tang

While reinforcement learning (RL) can empower autonomous agents by enabling self-improvement through interaction, its practical adoption remains challenging due to costly rollouts, limited task diversity, unreliable reward signals, and…

CybORG++ is an advanced toolkit for reinforcement learning research focused on network defence. Building on the CAGE 2 CybORG environment, it introduces key improvements, including enhanced debugging capabilities, refined agent…

密码学与安全 · 计算机科学 2024-10-23 Harry Emerson , Liz Bates , Chris Hicks , Vasilios Mavroudis

The rapid increase in the number of cyber-attacks in recent years raises the need for principled methods for defending networks against malicious actors. Deep reinforcement learning (DRL) has emerged as a promising approach for mitigating…

机器学习 · 计算机科学 2024-09-30 Gregory Palmer , Chris Parry , Daniel J. B. Harrold , Chris Willis

Recent advances in large language models (LLMs) have sparked growing interest in building generalist agents that can learn through online interactions. However, applying reinforcement learning (RL) to train LLM agents in multi-turn,…

Autonomous Cyber Defence is required to respond to high-tempo cyber-attacks. To facilitate the research in this challenging area, we explore the utility of the autonomous cyber operation environments presented as part of the Cyber Autonomy…

密码学与安全 · 计算机科学 2023-09-15 Mitchell Kiely , David Bowman , Maxwell Standen , Christopher Moir

Existing reinforcement learning environment libraries use monolithic environment classes, provide shallow methods for altering agent observation and action spaces, and/or are tied to a specific simulation environment. The Core Reinforcement…

Integrating LLM and reinforcement learning (RL) agent effectively to achieve complementary performance is critical in high stake tasks like cybersecurity operations. In this study, we introduce SecurityBot, a LLM agent mentored by…

密码学与安全 · 计算机科学 2024-03-27 Yikuan Yan , Yaolun Zhang , Keman Huang

Reinforcement Learning (RL) agents demonstrating proficiency in a training environment exhibit vulnerability to adversarial perturbations in input observations during deployment. This underscores the importance of building a robust agent…

机器学习 · 计算机科学 2024-08-05 Tung M. Luu , Haeyong Kang , Tri Ton , Thanh Nguyen , Chang D. Yoo

Safe interaction with the environment is one of the most challenging aspects of Reinforcement Learning (RL) when applied to real-world problems. This is particularly important when unsafe actions have a high or irreversible negative impact…

机器学习 · 计算机科学 2021-10-22 Erik Aumayr , Saman Feghhi , Filippo Vannella , Ezeddin Al Hakim , Grigorios Iakovidis
‹ 上一页 1 2 3 10 下一页 ›