Skip to content

Author

Ravi Tandon

We have 2 of 14 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Oct 2026

Intent-Hiding Jailbreaks: An Information-Theoretic Framework for Compositional Attacks

Recent work has shown that large language models (LLMs) can be vulnerable to jailbreak attacks in which harmful intent is obscured through composition with benign tasks. A harmful request refused in isolation may elicit a different response when embedded within a larger, seemingly benign query. We study these compositi...

Feng-Wei Tian, Ravi Tandon · 0 citations
#machine learning Preprint Sep 2026

Adaptive Multi-Value Control in LLMs via Causal Activation Steering

Large language models (LLMs) are increasingly deployed in settings where responses must reflect multiple, potentially interacting social norms and human values. Activation steering offers a lightweight alternative to training-based alignment by modifying internal activations at inference time. However, prior human-valu...

Payel Bhattacharjee, Ravi Tandon · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.