Skip to content

Author

Dylan Zhang

We have 2 of 21 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Selecting Diverse SFT Traces Improves Post-RL Generalization

Verified solutions are not equally useful for preparing reasoning models for reinforcement learning (RL). We present a comprehensive study of route diversity, the variation in the sequences of reasoning steps in supervised fine-tuning (SFT) data, and propose a lightweight, rule-based fingerprint to select for it. From...

Dylan Zhang, Ming-Yuan Wu, Jin-Ning Li · 0 citations

Useful Memories Become Faulty When Continuously Updated by LLMs

This work traces the regression to the consolidation step rather than the underlying experience: the same trajectories yield qualitatively different memories under different update schedules, and an episodic-only control that simply retains those trajectories remains competitive with the consolidators the authors test.

Dylan Zhang, Yan-Shan Lin, Zheng Wu et al. · 10 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.