Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Learning to Steer, Steering to See: Unveiling the Geometry of RLVR in Large Language Models via Trainable Vectors

The Alpha-Stabler framework is proposed, a plug-and-play framework with a Predictor that monitors principal-subspace intrusion for early collapse warnings, and a Controller that removes the principal-subspace component of activation gradients during backpropagation while preserving the orthogonal complement.

Yu-Chen Cai, Ding Cao, Qi-Xiang Yin et al. · 1 citation
#artificial intelligence Preprint Sep 2026

SAGE: Structured Strategic Reasoning for Efficient LLM Game Playing

A strong LLM strategic agent should reason prospectively over uncertain futures, adapt its strategy to opponents'behavioral tendencies, and continuously recalibrate its decision process from interaction experience. However, incorporating these sources in free-form reasoning could lead to unsupported strategic assumptio...

Zhi-Wei Chen, Tian-Chun Wang, Zhong-Tao Rao et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.