Skip to content

Author

Sarthak Mittal

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Amortizing intractable inference in diffusion models for vision, language, and control

Amortized sampling of the posterior over data is studied, and the asymptotic correctness of a data-free learning objective, relative trajectory balance, is proved for training a diffusion model that samples from this posterior, a problem that existing methods solve only approximately or in restricted cases.

S. Venkatraman, Moksh Jain, Luca Scimeca et al. · 75 citations · ⚡5

Improved off-policy training of diffusion samplers

This work benchmarks several diffusion-structured inference methods, including simulation-based variational approaches and off-policy methods (continuous generative flow networks), and proposes a novel exploration strategy for off-policy methods, based on local search in the target space with the use of a replay buffer.

Marcin Sendera, Minsu Kim, Sarthak Mittal et al. · 52 citations · ⚡7