#machine learning
Feb 2024
Improved off-policy training of diffusion samplers
This work benchmarks several diffusion-structured inference methods, including simulation-based variational approaches and off-policy methods (continuous generative flow networks), and proposes a novel exploration strategy for off-policy methods, based on local search in the target space with the use of a replay buffer.
Marcin Sendera, Minsu Kim, Sarthak Mittal et al.
· Neural Information Processin... · 52 citations
· ⚡7