Skip to content

Author

Zi-Heng Cheng

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Oct 2026

How RL Reshapes LLM Reasoning: Transferability, Coverage, and Scaling Laws

Recent studies on reinforcement learning (RL) report seemingly conflicting evidence about large language model (LLM) reasoning. Training on mathematics can improve performance in other domains, yet gains in Pass@1 can coincide with lower Pass@$N$ than the base model. This raises a fundamental question: does RL expand a...

Zi-Heng Cheng, Yi-Xiao Huang, Han-Lin Zhu et al. · 0 citations
#machine learning Preprint Sep 2026

Aligning One-Step Generative Models with Reward-Weighted Transport Distillation

Theoretical analysis shows that the fixed-point distributions of RWTD interpolate between off-policy reward tilting of the reference and on-policy tilting of the current model, providing a principled approach to balancing reward adaptation with retention of prior knowledge.

Austin S. Wang, Zi-Heng Cheng, Le-Xing Ying · 0 citations

Multi-Mask Diffusion Language Models for Few-Step Generation

This work proposes a multi-mask diffusion model (MultiMDM) that preserves the masking structure towards few-step generation and derives a closed-form ELBO training objective for MultiMDM that supports continual training from pretrained MDMs.

Sijin Chen, Yinuo Ren, Heyang Zhao et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.