Preprint
Jul 2026
HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion
Experiments show that HVA establishes a new Pareto frontier for training-free sparse attention in video diffusion, reducing end-to-end latency by up to $2.13\times while improving fidelity over existing training-free sparse attention baselines.
Dongyeun Lee, A. Zandieh, V. Mirrokni et al.
· 1 citation