Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

N_0-VTLA: Scaling Vision-Tactile-Language-Action Models with Latent Tactile Tokens

First seen · 8/3/2026, 12:00 PMLatest activity · 8/3/2026, 12:00 PM

The paper introduces N_0-VTLA, a vision-tactile-language-action foundation model for contact-rich manipulation and offline policy improvement. Its training recipe combines visuo-tactile pretraining on the NeoData robot dataset, staged tactile-pathway integration, and ALTER, an advantage-conditioned offline reinforcement learning method. The authors report wins on all nine NeoReal real-robot tasks and 63.8% mean success across a 20-task simulation suite, compared with 44.0% for the strongest baseline. With ALTER, the policy reaches 75–95% success on three long-horizon real-robot tasks. The results are promising, but the supplied evidence is limited to the paper abstract.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers8/3, 12:00 PMnot independentRepresentative
    N_0-VTLA: Scaling Vision-Tactile-Language-Action Models with Latent Tactile Tokens