Computer Vision and Pattern Recognition · Computer Science
Adversarial Prompt Injection Attack on Multimodal Large Language Models
Meiwen Ding, Song Xia, Chenqi Kong, Xudong Jiang
2026-04-01
Computer Vision and Pattern Recognition · Computer Science
Image-based Prompt Injection: Hijacking Multimodal LLMs through Visually Embedded Adversarial Instructions
Neha Nagaraja, Lan Zhang, Zhilong Wang, Bo Zhang +1
2026-03-05
Cryptography and Security · Computer Science
Self-interpreting Adversarial Images
Tingwei Zhang, Collin Zhang, John X. Morris, Eugene Bagdasarian +1
2025-06-16
Cryptography and Security · Computer Science
Misusing Tools in Large Language Models With Visual Adversarial Examples
Xiaohan Fu, Zihan Wang, Shuheng Li, Rajesh K. Gupta +3
2023-10-06
Cryptography and Security · Computer Science
Con Instruction: Universal Jailbreaking of Multimodal Large Language Models via Non-Textual Modalities
Jiahui Geng, Thy Thy Tran, Preslav Nakov, Iryna Gurevych
2025-06-03
Cryptography and Security · Computer Science
Prompt-in-Content Attacks: Exploiting Uploaded Inputs to Hijack LLM Behavior
Zhuotao Lian, Weiyu Wang, Qingkui Zeng, Toru Nakanishi +2
2025-08-28
Cryptography and Security · Computer Science
Defending Against Indirect Prompt Injection Attacks With Spotlighting
Keegan Hines, Gary Lopez, Matthew Hall, Federico Zarfati +2
2024-03-25
Machine Learning · Computer Science
Hijacking Large Language Models via Adversarial In-Context Learning
Xiangyu Zhou, Yao Qiang, Saleh Zare Zade, Prashant Khanduri +1
2025-05-30
Cryptography and Security · Computer Science
Misleading Large Language Models used (or misused) in Scientific Peer-Reviewing via Hidden Prompt-Injection Attacks
Matteo Gioele Collu, Umberto Salviati, Roberto Confalonieri, Mauro Conti +1
2026-03-31
Computer Vision and Pattern Recognition · Computer Science
InstructTA: Instruction-Tuned Targeted Attack for Large Vision-Language Models
Xunguang Wang, Zhenlan Ji, Pingchuan Ma, Zongjie Li +1
2024-06-27
Computer Vision and Pattern Recognition · Computer Science
Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection
Zedian Shao, Hongbin Liu, Yuepeng Hu, Neil Zhenqiang Gong
2026-04-13
Cryptography and Security · Computer Science
Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection
Kai Greshake, Sahar Abdelnabi, Shailesh Mishra, Christoph Endres +2
2023-05-08
Artificial Intelligence · Computer Science
Universal Adversarial Attack on Aligned Multimodal LLMs
Temurbek Rahmatullaev, Polina Druzhinina, Nikita Kurdiukov, Matvey Mikhalchuk +2
2025-06-06
Computation and Language · Computer Science
PandaGPT: One Model To Instruction-Follow Them All
Yixuan Su, Tian Lan, Huayang Li, Jialu Xu +2
2023-05-29
Cryptography and Security · Computer Science
Misaligned Roles, Misplaced Images: Structural Input Perturbations Expose Multimodal Alignment Blind Spots
Erfan Shayegani, G M Shahariar, Sara Abdali, Lei Yu +2
2025-04-08
Cryptography and Security · Computer Science
Breaking the Prompt Wall (I): A Real-World Case Study of Attacking ChatGPT via Lightweight Prompt Injection
Xiangyu Chang, Guang Dai, Hao Di, Haishan Ye
2025-04-24
Computation and Language · Computer Science
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
Brennen Hill, Surendra Parla, Venkata Abhijeeth Balabhadruni, Atharv Prajod Padmalayam +1
2025-09-08
Cryptography and Security · Computer Science
Adversarial Illusions in Multi-Modal Embeddings
Tingwei Zhang, Rishi Jha, Eugene Bagdasaryan, Vitaly Shmatikov
2025-08-26
Cryptography and Security · Computer Science
An LLM can Fool Itself: A Prompt-Based Adversarial Attack
Xilie Xu, Keyi Kong, Ning Liu, Lizhen Cui +3
2023-10-23