Computation and Language · Computer Science
LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers
Lingyao Li, Junjie Xiong, Changjia Zhu, Runlong Yu +4
2026-05-26
Cryptography and Security · Computer Science
Is Your Prompt Safe? Investigating Prompt Injection Attacks Against Open-Source LLMs
Jiawen Wang, Pritha Gupta, Ivan Habernal, Eyke Hüllermeier
2025-05-21
Artificial Intelligence · Computer Science
Automatic and Universal Prompt Injection Attacks against Large Language Models
Xiaogeng Liu, Zhiyuan Yu, Yizhe Zhang, Ning Zhang +1
2024-03-11
Cryptography and Security · Computer Science
Misleading Large Language Models used (or misused) in Scientific Peer-Reviewing via Hidden Prompt-Injection Attacks
Matteo Gioele Collu, Umberto Salviati, Roberto Confalonieri, Mauro Conti +1
2026-03-31
Cryptography and Security · Computer Science
A Critical Evaluation of Defenses against Prompt Injection Attacks
Yuqi Jia, Zedian Shao, Yupei Liu, Jinyuan Jia +2
2025-05-27
Cryptography and Security · Computer Science
Systematically Analyzing Prompt Injection Vulnerabilities in Diverse LLM Architectures
Victoria Benjamin, Emily Braca, Israel Carter, Hafsa Kanchwala +10
2024-11-01
Cryptography and Security · Computer Science
Ignore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking Competition
Sander Schulhoff, Jeremy Pinto, Anaum Khan, Louis-François Bouchard +6
2024-03-05
Computation and Language · Computer Science
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
Mansi Phute, Alec Helbling, Matthew Hull, ShengYun Peng +3
2024-05-03
Cryptography and Security · Computer Science
LLMs Cannot Reliably Judge (Yet?): A Comprehensive Assessment on the Robustness of LLM-as-a-Judge
Songze Li, Chuokun Xu, Jiaying Wang, Xueluan Gong +5
2025-11-18
Cryptography and Security · Computer Science
Prompt Injection Attacks on Large Language Models in Oncology
Jan Clusmann, Dyke Ferber, Isabella C. Wiest, Carolin V. Schneider +4
2025-03-20
Cryptography and Security · Computer Science
How Vulnerable Are AI Agents to Indirect Prompt Injections? Insights from a Large-Scale Public Competition
Mateusz Dziemian, Maxwell Lin, Xiaohan Fu, Micha Nowak +27
2026-03-18
Computation and Language · Computer Science
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
Brennen Hill, Surendra Parla, Venkata Abhijeeth Balabhadruni, Atharv Prajod Padmalayam +1
2025-09-08
Computers and Society · Computer Science
Prompt Injection Vulnerability of Consensus Generating Applications in Digital Democracy
Jairo Gudiño-Rosero, Clément Contet, Umberto Grandi, César A. Hidalgo
2026-03-03
Cryptography and Security · Computer Science
A Novel Evaluation Framework for Assessing Resilience Against Prompt Injection Attacks in Large Language Models
Daniel Wankit Yip, Aysan Esmradi, Chun Fai Chan
2024-01-03
Cryptography and Security · Computer Science
An LLM can Fool Itself: A Prompt-Based Adversarial Attack
Xilie Xu, Keyi Kong, Ning Liu, Lizhen Cui +3
2023-10-23
Cryptography and Security · Computer Science
Too Easily Fooled? Prompt Injection Breaks LLMs on Frustratingly Simple Multiple-Choice Questions
Xuyang Guo, Zekai Huang, Zhao Song, Jiahao Zhang
2025-08-20
Cryptography and Security · Computer Science
Defending Against Indirect Prompt Injection Attacks With Spotlighting
Keegan Hines, Gary Lopez, Matthew Hall, Federico Zarfati +2
2024-03-25
Cryptography and Security · Computer Science
Evaluation of Prompt Injection Defenses in Large Language Models
Priyal Deep, Shane Emmons, Amy Fox, Kyle Bacon +3
2026-05-14