Skip to content

Author

Chen-Hang Cui

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

EmoRSS: Mitigating Emotion-Induced Over-Refusal in Large Language Models

Emotional expression can influence the safety decisions of large language models (LLMs), offering a potential avenue for improving safety alignment. Existing studies have mainly focused on how emotional expressions facilitate attacks under harmful requests, while overlooking their effects on benign requests. We find th...

Shu-Yi Miao, Yao-Jin Ma, Chen-Hang Cui et al. · 0 citations
#artificial intelligence Preprint Sep 2026

ACTR: Aligning Thoughts and Responses for Multilingual Safety in Reasoning LLMs

This work proposes aligning cross-lingual thoughts and responses (ACTR), a framework that improves multilingual safety alignment by strengthening the use of existing safety reasoning, and devise neuron-selective consistency optimization (NSCO), which uses a frozen judge model to reward agreement between the safety cate...

Xian-Hui Zhang, Jian Yu, Cheng-Yu Xie et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.