This is the Replicated Computational Results (RCR) Report for the paper ``Can LLMs Hack Enterprise Networks?" The paper empirically investigates the efficacy and effectiveness of different LLMs for penetration-testing enterprise networks, i.e., Microsoft Active Directory Assumed-Breach Simulations. This RCR report describes the artifacts used in the paper, how to create an evaluation setup, and highlights the analysis scripts provided within our prototype.
Cite
@article{arxiv.2603.01789,
title = {Can LLMs Hack Enterprise Networks? -- Replicated Computational Results (RCR) Report},
author = {Andreas Happe and Jürgen Cito},
journal= {arXiv preprint arXiv:2603.01789},
year = {2026}
}