Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

ELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training Transformers

First seen · 7/22/2026, 10:34 PMLatest activity · 7/22/2026, 10:34 PM

ELSAA approximates the attention score operator itself after dense Q, K, and V projections, rather than factorizing the Transformer’s learned projection or output matrices. It combines a sparse branch for selected high-similarity interactions with a low-rank branch for diffuse global context. Because the two branches may have substantially different normalization mass, ELSAA adds a denominator-aware fusion term to rescale the sparse contribution. The stated goal is longer-context Transformer training without materializing the full quadratic attention matrix.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/22, 10:34 PMnot independentRepresentative
    ELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training Transformers