Cryptography and Security · Computer Science
Evaluation of Prompt Injection Defenses in Large Language Models
Priyal Deep, Shane Emmons, Amy Fox, Kyle Bacon +3
2026-05-14
Cryptography and Security · Computer Science
PACEbench: A Framework for Evaluating Practical AI Cyber-Exploitation Capabilities
Zicheng Liu, Lige Huang, Jie Zhang, Dongrui Liu +2
2025-10-14
Machine Learning · Computer Science
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks
Yifan Zeng, Yiran Wu, Xiao Zhang, Huazheng Wang +1
2024-11-15
Cryptography and Security · Computer Science
AutoAttacker: A Large Language Model Guided System to Implement Automatic Cyber-attacks
Jiacen Xu, Jack W. Stokes, Geoff McDonald, Xuesong Bai +4
2024-03-05
Networking and Internet Architecture · Computer Science
Forewarned is Forearmed: A Survey on Large Language Model-based Agents in Autonomous Cyberattacks
Minrui Xu, Jiani Fan, Xinyu Huang, Conghao Zhou +7
2025-05-28
Cryptography and Security · Computer Science
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
Qianshan Wei, Tengchao Yang, Yaochen Wang, Xinfeng Li +6
2025-10-06
Cryptography and Security · Computer Science
Your Agent Can Defend Itself against Backdoor Attacks
Li Changjiang, Liang Jiacheng, Cao Bochuan, Chen Jinghui +1
2025-06-12
Cryptography and Security · Computer Science
AgentSecBench: Measuring Prompt Injection, Privacy Leakage, and Tool-Use Integrity in LLM Agents
Faruk Alpay, Taylan Alpay
2026-05-27
Cryptography and Security · Computer Science
WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents
Yulin Chen, Tri Cao, Haoran Li, Yue Liu +6
2026-04-15
Cryptography and Security · Computer Science
Automating Security Audit Using Large Language Model based Agent: An Exploration Experiment
Jia Hui Chin, Pu Zhang, Yu Xin Cheong, Jonathan Pan
2025-05-19
Cryptography and Security · Computer Science
Large Language Models in Cybersecurity: State-of-the-Art
Farzad Nourmohammadzadeh Motlagh, Mehrdad Hajizadeh, Mehryar Majd, Pejman Najafi +2
2024-02-05
Cryptography and Security · Computer Science
Autonomous Adversary: Red-Teaming in the age of LLM
Mohammad Mamun, Mohamed Gaber, Scott Buffett, Sherif Saad
2026-05-08
Multiagent Systems · Computer Science
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
Rana Muhammad Shahroz Khan, Zhen Tan, Sukwon Yun, Charles Fleming +1
2025-10-10
Cryptography and Security · Computer Science
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner
Xunguang Wang, Daoyuan Wu, Zhenlan Ji, Zongjie Li +6
2025-02-06
Cryptography and Security · Computer Science
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
Wei Zhao, Zhe Li, Peixin Zhang, Jun Sun
2026-05-12
Cryptography and Security · Computer Science
Construction and Evaluation of LLM-based agents for Semi-Autonomous penetration testing
Masaya Kobayashi, Masane Fuchi, Amar Zanashir, Tomonori Yoneda +1
2025-02-24
Cryptography and Security · Computer Science
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification
Boyang Zhang, Yicong Tan, Yun Shen, Ahmed Salem +3
2024-07-31
Cryptography and Security · Computer Science
Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels
Chenghao Du, Quanfeng Huang, Tingxuan Tang, Zihao Wang +2
2025-11-07
Cryptography and Security · Computer Science
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
Yihan Wang, Huanqi Yang, Shantanu Pal, Weitao Xu
2025-12-25
Cryptography and Security · Computer Science
Imprompter: Tricking LLM Agents into Improper Tool Use
Xiaohan Fu, Shuheng Li, Zihan Wang, Yihao Liu +3
2024-10-23
Cryptography and Security · Computer Science
AgentSentinel: An End-to-End and Real-Time Security Defense Framework for Computer-Use Agents
Haitao Hu, Peng Chen, Yanpeng Zhao, Yuqi Chen
2025-09-10
Cryptography and Security · Computer Science
ZeroDayBench: Evaluating LLM Agents on Unseen Zero-Day Vulnerabilities for Cyberdefense
Nancy Lau, Louis Sloot, Jyoutir Raj, Giuseppe Marco Boscardin +5
2026-03-04
Artificial Intelligence · Computer Science
Large Language Models are Autonomous Cyber Defenders
Sebastián R. Castro, Roberto Campbell, Nancy Lau, Octavio Villalobos +2
2025-07-22