Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression

First seen · 7/27/2026, 12:00 PMLatest activity · 7/27/2026, 12:00 PM

VisCo introduces a training-efficient self-compression framework that reuses a pretrained vision-language model as its own visual token compressor. It uses a small set of memory tokens in a parameter-sharing autoencoder and transfers hierarchical information from encoding to decoding. According to the paper’s abstract, VisCo outperforms prior methods across all evaluated compression ratios, with larger gains at aggressive compression levels, and remains stable even when compressing to a single token. Combining the learned memory tokens with the original visual tokens can also improve the base model, suggesting that the compressed representation may contain complementary information.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/27, 12:00 PMnot independentRepresentative
    VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression