Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

TORINO: Token Reduction via Interpretable Concept Overlap in Vision-Language Models

First seen · 7/6/2026, 09:43 AMLatest activity · 7/6/2026, 09:43 AM

TORINO is a plug-and-play framework for reducing visual tokens in vision-language models without fine-tuning the underlying model. It uses sparse autoencoders (SAEs) to map visual tokens into an interpretable latent space, groups tokens by shared concept activations, and applies pruning or merging within each group. The method dynamically adapts the reduction rate to image complexity instead of enforcing a fixed token budget. The abstract reports favorable efficiency-accuracy trade-offs across multiple VLM benchmarks, but provides no concrete compression ratios, latency measurements, or benchmark results.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/6, 09:43 AMnot independentRepresentative
    TORINO: Token Reduction via Interpretable Concept Overlap in Vision-Language Models