
Partition the Support, Reconstruct the Residual: Training-Free Sparse Attention for Video Generation and World Models
arXiv, 2026 · Under review
Approximates dense attention with sparse computation and residual correction, accelerating video generators and world models without retraining. Achieves 1.48–2.61× end-to-end speedups while preserving generation quality.





