Skip to content

Author

Philipp Krähenbühl

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Jul 2026

Mask-Aware Policy Gradients for Diffusion Language Models

This work observes that MDLM generation involves two decisions at each step: what tokens to place at each masked position and which positions to remask, and formalizes this as a two-stage action MDP, showing that the policy gradient naturally decomposes into a token term and a masking term.

Haran Raajesh, Kulin Shah, Adam R. Klivans et al. · 1 citation