Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingNewsWatching0 independent reports0

ABot-N1: Toward a General Visual Language Navigation Foundation Model

First seen · 7/14/2026, 12:00 PMLatest activity · 7/14/2026, 12:00 PM

ABot-N1 proposes a slow-fast architecture for general visual-language navigation. A slow vision-language reasoner produces explicit linguistic reasoning and a pixel-space goal, while a fast action expert uses those signals to generate continuous waypoints at the native control frequency. The pixel anchors act as a shared interface across point-goal, object-goal, POI-goal, instruction-following, and person-following tasks. According to the supplied abstract, ABot-N1 improves POI arrival by 35.0 percentage points to 77.3%, reaches 95.4% and 92.9% success rates in complex indoor and outdoor scenes, and releases new Point-Goal and POI-Goal benchmarks.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/12, 12:21 AMnot independent
    ABot-N1: Toward a General Visual Language Navigation Foundation Model
  2. AggregatorHuggingFace Daily Papers7/14, 12:00 PMnot independentRepresentative
    ABot-N1: Toward a General Visual Language Navigation Foundation Model