Switch language한국어
Back to the list

DFSAttn: Dynamic Fine-grained Sparse Attention for Efficient Video Generation

TL;DR AI

Key summary

2 min read
  1. Researchers introduced DFSAttn, a training-free sparse attention method for diffusion transformers in video generation.

  2. DFSAttn combines token reordering, hierarchical block scoring, and sparse mask caching with adaptive ratios for fine-grained sparsity.

  3. The method cuts the high cost of full spatiotemporal attention and delivers up to 2.1x end-to-end speedup.

  4. It also maintains or improves quality versus prior sparse-attention methods, especially at high sparsity levels.

Read the original