Skip to content

Author

Jian Yu

We have 2 of 8 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

ACTR: Aligning Thoughts and Responses for Multilingual Safety in Reasoning LLMs

This work proposes aligning cross-lingual thoughts and responses (ACTR), a framework that improves multilingual safety alignment by strengthening the use of existing safety reasoning, and devise neuron-selective consistency optimization (NSCO), which uses a frozen judge model to reward agreement between the safety cate...

Xian-Hui Zhang, Jian Yu, Cheng-Yu Xie et al. · 0 citations
Jul 2026

SafeNexus: Discovering and Steering Modality-Universal Safety Neurons in MLLMs

SafeNexus is introduced, a cross-modal safety alignment framework that adopts a dedicated neuron-level intervention strategy that outperforms prevailing state-of-the-art approaches on safety benchmarks spanning diverse modality combinations, while effectively preserving utility.

Jian Yu, Fei Shen, Cong Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.