Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Masked Visual Actions for Unified World Modeling

First seen · 7/22/2026, 12:00 PMLatest activity · 7/22/2026, 12:00 PM

This paper introduces Masked Visual Actions, a pixel-space control interface that represents actions as partially revealed trajectories of arbitrary entities in a video. Revealing robot motion turns a video model into a forward dynamics predictor for scene responses, while revealing desired object motion enables inverse modeling of robot behavior. A single checkpoint is fine-tuned on 15 hours of masked examples from real videos and simulation. The authors report strong visual fidelity and controllability across diverse scenes and embodiments, plus utility for imagined rollouts, model-based planning, and manipulation policy evaluation.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/22, 12:00 PMnot independentRepresentative
    Masked Visual Actions for Unified World Modeling