KV cache compression is widely used for long context LLM inference under memory constraints, while deployed systems typically score refusals after generation with keyword filters or learned classifiers. Such monitors are intended to indicate whether a model declined a harmful request under the serving regime actually u...
Kang Chen, Xiu-Ze Zhou, Hong Chen et al.· 0 citations
A novel neural network called the Co-occurrence Graph Neural Network (CoGNN), which utilizes two co-occurrence graphs to establish user and item relationships and outperforms various baseline models in terms of recommendation accuracy and algorithm convergence.
Chao Lin, Y. Lin, You-Yu Wang et al.· Multimedia Systems· 0 citations
This survey offers a comprehensive overview of the main data security risks facing LLMs and reviews current defense strategies, including adversarial training, data cleaning, output guardrails, Reinforcement Learning from Human Feedback, data augmentation, and Retrieval-Augmented Generation (RAG)/agent defenses.
Kang Chen, Xiuze Zhou, Yuanhui Yu et al.· Journal of King Saud Univers...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.