Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Continue or Replan? Bernoulli-Continuation Policy Learning for Adaptive Horizon Execution

First seen · 8/4/2026, 07:21 PMLatest activity · 8/4/2026, 07:21 PM

This paper introduces Bernoulli-Continuation Policy (BCP), a plug-and-play mechanism that adaptively decides whether a frozen Vision-Language-Action model should continue executing its current action chunk or replan. Its continuation head is trained through trajectory-level reinforcement learning with a reward balancing task success and replanning efficiency. The abstract reports gains on RoboTwin 2.0, LIBERO, LIBERO-PRO, and two real-robot manipulation tasks. RoboTwin success rises from 89.88% to 93.94% across 50 tasks, while real-robot success improves from 74% to 92% and from 44% to 84%.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv8/4, 07:21 PMnot independentRepresentative
    Continue or Replan? Bernoulli-Continuation Policy Learning for Adaptive Horizon Execution