Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

A Deep Reinforcement Learning Algorithm for the Vehicle Routing Problem with Stochastic Demands and Outsourcing

First seen · 7/19/2026, 12:27 AMLatest activity · 7/19/2026, 12:27 AM

This paper studies the vehicle routing problem with stochastic demands and outsourcing (VRP-SDO). It partitions customers between a fixed fleet and a common carrier, then estimates the dynamic routing cost of the committed subset with an offline-trained deep Q-network. The policy uses a graph attention network to aggregate customer and vehicle information for each acting vehicle. The authors report a 19.6% routing-cost reduction versus a state-of-the-art method, at least 29.6% versus classical heuristics, and a 13.7% improvement over a non-attention variant. Decisions are generated within minutes, while benchmarks without an offline estimator take over an hour.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/19, 12:27 AMnot independentRepresentative
    A Deep Reinforcement Learning Algorithm for the Vehicle Routing Problem with Stochastic Demands and Outsourcing