Preprint
Sep 2026
Parallel Policy-Gradient Methods for Parameter Optimization of Nonlinear Feedback Controllers
This letter develops a time-parallel policy-gradient framework for discrete-time nonlinear control-affine systems and proves that, for any finite horizon T, the state solver recovers the exact trajectory from any initialization in at most T iterations.
A. Nguyen, Leilei Cui
· 0 citations