Cryptography and Security · Computer Science
Safeguarding Large Language Models: A Survey
Yi Dong, Ronghui Mu, Yanghao Zhang, Siqi Sun +8
2024-06-06
Cryptography and Security · Computer Science
Guardrails for trust, safety, and ethical development and deployment of Large Language Models (LLM)
Anjanava Biswas, Wrick Talukdar
2026-01-22
Cryptography and Security · Computer Science
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly
Yifan Yao, Jinhao Duan, Kaidi Xu, Yuanfang Cai +2
2024-03-22
Computation and Language · Computer Science
MrGuard: A Multilingual Reasoning Guardrail for Universal LLM Safety
Yahan Yang, Soham Dan, Shuo Li, Dan Roth +1
2025-09-29
Cryptography and Security · Computer Science
Large Language Model Supply Chain: Open Problems From the Security Perspective
Qiang Hu, Xiaofei Xie, Sen Chen, Lei Ma
2024-11-05
Cryptography and Security · Computer Science
ChatGPT and Other Large Language Models for Cybersecurity of Smart Grid Applications
Aydin Zaboli, Seong Lok Choi, Tai-Jin Song, Junho Hong
2024-02-27
Cryptography and Security · Computer Science
Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models
GodsGift Uzor, Hasan Al-Qudah, Ynes Ineza, Abdul Serwadda
2025-08-26
Computation and Language · Computer Science
TrafficSafetyGPT: Tuning a Pre-trained Large Language Model to a Domain-Specific Expert in Transportation Safety
Ou Zheng, Mohamed Abdel-Aty, Dongdong Wang, Chenzhu Wang +1
2023-07-31
Machine Learning · Computer Science
SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks
Xiangman Li, Xiaodong Wu, Qi Li, Jianbing Ni +1
2025-08-22
Computation and Language · Computer Science
SafetyPrompts: a Systematic Review of Open Datasets for Evaluating and Improving Large Language Model Safety
Paul Röttger, Fabio Pernisi, Bertie Vidgen, Dirk Hovy
2025-01-13
Computation and Language · Computer Science
Self-Guard: Empower the LLM to Safeguard Itself
Zezhong Wang, Fangkai Yang, Lu Wang, Pu Zhao +4
2024-03-25
Cryptography and Security · Computer Science
On Protecting the Data Privacy of Large Language Models (LLMs): A Survey
Biwei Yan, Kun Li, Minghui Xu, Yueyan Dong +3
2024-03-15
Cryptography and Security · Computer Science
Securing Large Language Models: Threats, Vulnerabilities and Responsible Practices
Sara Abdali, Richard Anarfi, CJ Barberan, Jia He +1
2025-06-13
Computation and Language · Computer Science
LLMGuard: Guarding Against Unsafe LLM Behavior
Shubh Goyal, Medha Hira, Shubham Mishra, Sukriti Goyal +5
2024-03-05
Computation and Language · Computer Science
When in Doubt, Cascade: Towards Building Efficient and Capable Guardrails
Manish Nagireddy, Inkit Padhi, Soumya Ghosh, Prasanna Sattigeri
2025-08-08
Computation and Language · Computer Science
Building Guardrails for Large Language Models
Yi Dong, Ronghui Mu, Gaojie Jin, Yi Qi +5
2024-05-30
Computation and Language · Computer Science
Multilingual Jailbreak Challenges in Large Language Models
Yue Deng, Wenxuan Zhang, Sinno Jialin Pan, Lidong Bing
2024-03-05
Computation and Language · Computer Science
The Need for Guardrails with Large Language Models in Medical Safety-Critical Settings: An Artificial Intelligence Application in the Pharmacovigilance Ecosystem
Joe B Hakim, Jeffery L Painter, Darmendra Ramcharran, Vijay Kara +5
2024-09-05