Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Dual-Stream Diffusion Net for Text-to-Video Generation

Aug 18, 2023

Binhui Liu, Xin Liu, Anbo Dai, Zhiyong Zeng, Zhen Cui, Jian Yang

Figure 1 for Dual-Stream Diffusion Net for Text-to-Video Generation

Figure 2 for Dual-Stream Diffusion Net for Text-to-Video Generation

Figure 3 for Dual-Stream Diffusion Net for Text-to-Video Generation

Figure 4 for Dual-Stream Diffusion Net for Text-to-Video Generation

Share this with someone who'll enjoy it:

Abstract:With the emerging diffusion models, recently, text-to-video generation has aroused increasing attention. But an important bottleneck therein is that generative videos often tend to carry some flickers and artifacts. In this work, we propose a dual-stream diffusion net (DSDN) to improve the consistency of content variations in generating videos. In particular, the designed two diffusion streams, video content and motion branches, could not only run separately in their private spaces for producing personalized video variations as well as content, but also be well-aligned between the content and motion domains through leveraging our designed cross-transformer interaction module, which would benefit the smoothness of generated videos. Besides, we also introduce motion decomposer and combiner to faciliate the operation on video motion. Qualitative and quantitative experiments demonstrate that our method could produce amazing continuous videos with fewer flickers.

* 8pages, 7 figures

View paper on

Share this with someone who'll enjoy it:

Title:Dual-Stream Diffusion Net for Text-to-Video Generation

Paper and Code