Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs

First seen · 8/4/2026, 12:00 PMLatest activity · 8/4/2026, 12:00 PM

This paper argues that exposing logged ground-truth future trajectories to teacher models during chain-of-thought annotation creates trajectory anchoring bias: models rationalize observed outcomes instead of inferring decisions from scene evidence, leading to less causally faithful reasoning and more severe hallucinations in difficult scenes. It introduces Autonomous-Driving Multiple-Choice Questions (AD-MCQ), which frames planning as selection among explicit candidate trajectories, and Deferred Exposure of Future Trajectories for RLVR (DEFT-RLVR), which withholds future trajectories until after the decision. The abstract reports improved autonomous-driving reasoning while preserving or improving general visual capabilities.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers8/4, 12:00 PMnot independentRepresentative
    Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs