Test-Time Alignment of Diffusion Models without reward over-optimization episode artwork

EPISODE · May 16, 2025 · 28 MIN

Test-Time Alignment of Diffusion Models without reward over-optimization

from Best AI papers explained · host Enoch H. Kang

This text introduces Diffusion Alignment as Sampling (DAS), a novel approach for aligning diffusion models with desired characteristics by treating the problem as sampling from a reward-aligned distribution. DAS utilizes a Sequential Monte Carlo (SMC) framework enhanced with tempering and a specially designed proposal distribution to efficiently generate high-reward samples without requiring additional training of the diffusion model. The method demonstrates superiority over existing guidance and fine-tuning techniques in single and multi-objective reward optimization, cross-reward generalization, diversity preservation, and online black-box optimization. Theoretical analysis supports the benefits of tempering for improving sample efficiency and mitigating issues like over-optimization and manifold deviation. Experiments across various tasks, including image generation with different reward functions and complex multimodal distributions, validate the practical effectiveness and broad applicability of DAS.

Episode metadata supplied by the publisher feed · Published May 16, 2025

Embed this episode

NOW PLAYING

Test-Time Alignment of Diffusion Models without reward over-optimization

0:00 28:01

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Best AI papers explained?

This episode is 28 minutes long.

When was this Best AI papers explained episode published?

This episode was published on May 16, 2025.

Can I download this Best AI papers explained episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!