Review
Sep 2026
Implicit Manipulation for Skill Selection in LLM Agents with Semantic Matching
This work identifies a new implicit attack surface for skill selection: even when the user prompt and skill description appear benign in isolation, their semantic relationship can still be strategically shaped to favor an attacker-chosen skill.
Qi-Kai Wang, Yong-Zhao Zhang, Zhi-Wei Chen et al.
· 1 citation