Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning

First seen · 8/4/2026, 12:00 PMLatest activity · 8/4/2026, 12:00 PM

The paper introduces the World Critic Model (WCM), a lightweight LeJEPA-based critic for vision-language-action reinforcement learning. Instead of estimating value from single-frame observations or backbone latents alone, WCM jointly predicts future latent states and values, providing an explicit world-modeling objective for temporal representation learning. It is designed to work with both on-policy and off-policy pipelines and is compatible with Pi0, Pi0.5, and OpenVLA-OFT. The authors report experiments across 149 tasks on four benchmarks, plus seven real-world manipulation tasks using OpenVLA-OFT and Pi0.5 with off-policy RL.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers8/4, 12:00 PMnot independentRepresentative
    WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning