Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers

First seen · 7/22/2026, 12:00 PMLatest activity · 7/22/2026, 12:00 PM

This paper introduces a causal interpretability framework for large diffusion transformers (DiTs), combining attention decomposition with interventions over token spans, heads, and layers. It finds that structural template tokens contain little prompt-specific information at the encoder output, yet become dominant image-to-text attention sinks and causally preserve object identity during denoising. Prompt semantics are first injected into image latents and then read back into template tokens. Based on this mechanism, the authors propose a training-free pruning rule that removes heads strongly attending to prompt tokens, reducing attention FLOPs by 20% with a 1.4-point GenEval drop.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/22, 12:00 PMnot independentRepresentative
    Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers