W EAVE : Efficient Co-Scheduling for Disaggregated RL Post-Training
Tianyuan Wu, Lunxi Cao, Yi-Chen Wei et al.
· 3 citations
2 papers indexed here
We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.
Not the right person? Other researchers publish under this name.
Rollplex is presented, a runtime that decomposes the reference and training phase and moves the prefix computation into the rollout decode window and achieves speedup over serial colocation and disaggregation under the same GPU budget, while preserving the synchronous RL update.