Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion

First seen · 7/3/2026, 02:46 PMLatest activity · 7/3/2026, 02:46 PM

HyperVAttention (HVA) is a training-free sparse-attention framework for Video Diffusion Transformers. It combines 3D local-window clustering, a hybrid update schedule that performs full clustering only at anchor denoising steps, and hardware-aware cluster merging designed around GPU CTA-aligned execution costs. The paper argues that these components reduce clustering overhead, avoid redundant cluster updates, and improve sparse block density. On text-to-video generation experiments, HVA reportedly reduces end-to-end latency by up to 2.13x while achieving higher fidelity than existing training-free sparse-attention baselines. The supplied abstract does not specify models, datasets, hardware, or detailed quality metrics.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/3, 02:46 PMnot independentRepresentative
    HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion