Preprint
Jul 2026
Improving LLM-Generated Process Model Quality Through Reinforcement Learning: The Role of Reward Function Design
It is demonstrated that reward composition is a primary determinant of optimization outcomes, with effects as large as the decision to apply RL itself, and generalize to any structured generation task where quality is assessed along multiple automated dimensions.
Alexander Rombach, Chantale Lauer, Nijat Mehdiyev
· 0 citations