arXiv · 2512.03234
Iterative Tilting for Diffusion Fine-Tuning
Abstract
We introduce iterative tilting, a gradient-free method for fine-tuning diffusion models toward reward-tilted distributions. The method decomposes a large reward tilt $\exp(\lambda r)$ into $N$ sequential smaller tilts, each admitting a tractable score update via first-order Taylor expansion. This requires only forward evaluations of the reward function and avoids backpropagating through sampling chains. We validate on a two-dimensional Gaussian mixture with linear reward, where the exact tilted distribution is available in closed form.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Jean Pachebat, Giovanni Conforti, Alain Durmus, Yazid Janati. 2025-12-02. Iterative Tilting for Diffusion Fine-Tuning. https://arxiv.org/abs/2512.03234
Cite the original work for its findings. Save a collection to share your selection of sources.