Skip to content

Author

Zhao-Jun Peng

4 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Bellman-Certified Rounding for Sparse Policy Deployment in MDPs

Continuous policy optimization may spread an update across many states, even when deployment permits only a few complete state-level changes. We study how much discounted return can be retained when continuous row mixtures are rounded to sparse binary policies in finite MDPs. Policy-dependent visitation couples the row...

Zhao-Jun Peng · 0 citations
#machine learning Preprint Sep 2026

Markovian Nonconvex ADMM for Reinforcement Learning: Bellman-Resolvent Stability Beyond Smooth Blocks

We identify and study a structural mechanism for Markovian nonconvex ADMM in reinforcement learning. Using finite discounted MDPs as a canonical proving ground, we show that the discounted Bellman resolvent $(I-\gamma P_\pi)^{-1}$ can provide the multiplier stability that classical nonconvex ADMM analyses often obtain...

Zhao-Jun Peng · 0 citations
#machine learning Preprint Sep 2026

Adapting to Decision-Relevant Non-Stationarity in Decentralized Heterogeneous Bandits

This work introduces Decision-Relevant Fresh Comparison (DRFC), which uses new, balanced samples from all agents to compare arms at the network level and switches only when fresh global evidence indicates that the common best arm has changed, and proves a high-probability dynamic regret bound with no adaptation term de...

Zhao-Jun Peng · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.