Review
A Survey on Actionable Interpretability in Large Language Models
This survey reviews LLM interpretability through the lens of actionability, presenting a taxonomy of attributional and mechanistic approaches, along with emerging methods tailored to vision–language models (VLMs), and examining how actionable interpretability supports downstream objectives.
Jie Cai, Mafizur Rahman, James Enouen et al.
· 0 citations