This survey offers a comprehensive overview of the main data security risks facing LLMs and reviews current defense strategies, including adversarial training, data cleaning, output guardrails, Reinforcement Learning from Human Feedback, data augmentation, and Retrieval-Augmented Generation (RAG)/agent defenses.
Kang Chen, Xiuze Zhou, Yuanhui Yu et al.· Journal of King Saud Univers...· 0 citations