Skip to content

Author

Feng-Wei Tian

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Oct 2026

Intent-Hiding Jailbreaks: An Information-Theoretic Framework for Compositional Attacks

Recent work has shown that large language models (LLMs) can be vulnerable to jailbreak attacks in which harmful intent is obscured through composition with benign tasks. A harmful request refused in isolation may elicit a different response when embedded within a larger, seemingly benign query. We study these compositi...

Feng-Wei Tian, Ravi Tandon · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.