Skip to content

Author

Ze-Ke Wang

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

HyperParallel-FSDP: Topology-Aware Fully Sharded Training with Layout-Driven Muon on Ascend SuperPods

Declarative SPMD programming uses tensor sharding descriptions to drive distributed execution, separating parallelization from model code. However, the evaluated PyTorch DTensor stack dispatches every operator below autograd, incurring repeated dispatch and metadata costs, while lacking an inexpensive end-to-end valida...

Mo Sun, Yi-Fan Yao, Yan-Wei Liu et al. · 0 citations
2026

SwCC-Sec: Secure Software-Programmable and Per-Packet Congestion Control in RDMA Engine

Many data centers adopt Remote Direct Memory Access (RDMA) to allow data center applications to achieve low latency and high throughput, while keeping minimal CPU overhead. The upper-layer applications keep evolving rapidly, and thus need congestion control algorithms (CCAs) that exist in the NIC hardware also to react...

Hong-Jing Huang, Xu-Zheng Chen, Zi-Yu Song et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Nereus: Adaptive Parallelism for LLM Post-Training

Nereus is a cost-aware runtime that adapts RL post-training jobs into efficient execution plans and executes a transition using a memory-feasible global plan and admits the transition using a cost model calibrated against the running job.

Songlin Jiang, Tuo Shi, Si-Tong Zhang et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.