Computation and Language · Computer Science
A Survey on Measuring and Mitigating Reasoning Shortcuts in Machine Reading Comprehension
Xanh Ho, Johannes Mario Meissner, Saku Sugawara, Akiko Aizawa
2023-09-07
Software Engineering · Computer Science
Debugging Without Error Messages: How LLM Prompting Strategy Affects Programming Error Explanation Effectiveness
Audrey Salmon, Katie Hammer, Eddie Antonio Santos, Brett A. Becker
2025-01-13
Cryptography and Security · Computer Science
Too Easily Fooled? Prompt Injection Breaks LLMs on Frustratingly Simple Multiple-Choice Questions
Xuyang Guo, Zekai Huang, Zhao Song, Jiahao Zhang
2025-08-20
Computation and Language · Computer Science
Navigating the Shortcut Maze: A Comprehensive Analysis of Shortcut Learning in Text Classification by Language Models
Yuqing Zhou, Ruixiang Tang, Ziyu Yao, Ziwei Zhu
2024-11-13
Computation and Language · Computer Science
Do LLMs Overcome Shortcut Learning? An Evaluation of Shortcut Challenges in Large Language Models
Yu Yuan, Lili Zhao, Kai Zhang, Guangting Zheng +1
2024-10-18
Computation and Language · Computer Science
Large Language Models as Misleading Assistants in Conversation
Betty Li Hou, Kejian Shi, Jason Phang, James Aung +2
2024-07-17
Computation and Language · Computer Science
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
Mansi Phute, Alec Helbling, Matthew Hull, ShengYun Peng +3
2024-05-03
Cryptography and Security · Computer Science
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification
Boyang Zhang, Yicong Tan, Yun Shen, Ahmed Salem +3
2024-07-31
Cryptography and Security · Computer Science
Misleading Large Language Models used (or misused) in Scientific Peer-Reviewing via Hidden Prompt-Injection Attacks
Matteo Gioele Collu, Umberto Salviati, Roberto Confalonieri, Mauro Conti +1
2026-03-31
Computation and Language · Computer Science
Shortcut Learning of Large Language Models in Natural Language Understanding
Mengnan Du, Fengxiang He, Na Zou, Dacheng Tao +1
2023-05-09
Computation and Language · Computer Science
GPT Editors, Not Authors: The Stylistic Footprint of LLMs in Academic Preprints
Soren DeHaan, Yuanze Liu, Johan Bollen, Sa'ul A. Blanco
2025-05-26
Computation and Language · Computer Science
Investigating Multi-Hop Factual Shortcuts in Knowledge Editing of Large Language Models
Tianjie Ju, Yijin Chen, Xinwei Yuan, Zhuosheng Zhang +3
2024-06-04
Artificial Intelligence · Computer Science
LLM Censorship: A Machine Learning Challenge or a Computer Security Problem?
David Glukhov, Ilia Shumailov, Yarin Gal, Nicolas Papernot +1
2023-07-25
Computer Vision and Pattern Recognition · Computer Science
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
Maan Qraitem, Nazia Tasnim, Piotr Teterwak, Kate Saenko +1
2025-02-14
Human-Computer Interaction · Computer Science
Detecting LLM-Generated Short Answers and Effects on Learner Performance
Shambhavi Bhushan, Danielle R Thomas, Conrad Borchers, Isha Raghuvanshi +4
2025-06-23
Artificial Intelligence · Computer Science
Protecting Publicly Available Data With Machine Learning Shortcuts
Nicolas M. Müller, Maximilian Burgert, Pascal Debus, Jennifer Williams +2
2023-10-31
Cryptography and Security · Computer Science
ChatGPT: Excellent Paper! Accept It. Editor: Imposter Found! Review Rejected
Kanchon Gharami, Sanjiv Kumar Sarkar, Safayat Bin Hakim, Yongxin Liu +2
2026-04-16
Computation and Language · Computer Science
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset
Owen Henkel, Libby Hills, Bill Roberts, Joshua McGrane
2024-05-07
Software Engineering · Computer Science
Understanding and supporting how developers prompt for LLM-powered code editing in practice
Daye Nam, Ahmed Omran, Ambar Murillo, Saksham Thakur +5
2025-12-22
Computation and Language · Computer Science
Shortcut Learning in In-Context Learning: A Survey
Rui Song, Yingji Li, Lida Shi, Fausto Giunchiglia +1
2024-12-02
Human-Computer Interaction · Computer Science
Decoding Logic Errors: A Comparative Study on Bug Detection by Students and Large Language Models
Stephen MacNeil, Paul Denny, Andrew Tran, Juho Leinonen +4
2023-11-28
Artificial Intelligence · Computer Science
Tool Preferences in Agentic LLMs are Unreliable
Kazem Faghih, Wenxiao Wang, Yize Cheng, Siddhant Bharti +4
2025-09-23