Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation

Mar 03, 2024

Hongjian Liu, Qingsong Xie, Zhijie Deng, Chen Chen, Shixiang Tang, Fueyang Fu, Zheng-jun Zha, Haonan Lu

Figure 1 for SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation

Figure 2 for SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation

Figure 3 for SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation

Figure 4 for SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation

Share this with someone who'll enjoy it:

Abstract:The iterative sampling procedure employed by diffusion models (DMs) often leads to significant inference latency. To address this, we propose Stochastic Consistency Distillation (SCott) to enable accelerated text-to-image generation, where high-quality generations can be achieved with just 1-2 sampling steps, and further improvements can be obtained by adding additional steps. In contrast to vanilla consistency distillation (CD) which distills the ordinary differential equation solvers-based sampling process of a pretrained teacher model into a student, SCott explores the possibility and validates the efficacy of integrating stochastic differential equation (SDE) solvers into CD to fully unleash the potential of the teacher. SCott is augmented with elaborate strategies to control the noise strength and sampling process of the SDE solver. An adversarial loss is further incorporated to strengthen the sample quality with rare sampling steps. Empirically, on the MSCOCO-2017 5K dataset with a Stable Diffusion-V1.5 teacher, SCott achieves an FID (Frechet Inceptio Distance) of 22.1, surpassing that (23.4) of the 1-step InstaFlow (Liu et al., 2023) and matching that of 4-step UFOGen (Xue et al., 2023b). Moreover, SCott can yield more diverse samples than other consistency models for high-resolution image generation (Luo et al., 2023a), with up to 16% improvement in a qualified metric. The code and checkpoints are coming soon.

* 22 pages, 16 figures

View paper on

Share this with someone who'll enjoy it:

Title:SCott: Accelerating Diffusion Models with Stochastic Consistency Distillation

Paper and Code