Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Hard Constraints, Smooth Gradients: Learning Feasible Inventory Policies via Differentiable Projection

First seen · 8/3/2026, 10:57 PMLatest activity · 8/3/2026, 10:57 PM

This paper embeds a differentiable convex optimization module into a deep reinforcement learning policy for constrained inventory control. A neural network proposes continuous action targets, a quadratic program projects them onto a relaxed feasible set, and a dual-informed integer mapping restores integrality while preserving feasibility. The method reports an average optimality gap below 1% on small instances, improvements of up to 9.75% over echelon base-stock policies and at least 7.7% over a rolling-horizon multistage stochastic program on larger networks, plus up to 3.22% cost reduction in an ASML industry-scale case study.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv8/3, 10:57 PMnot independentRepresentative
    Hard Constraints, Smooth Gradients: Learning Feasible Inventory Policies via Differentiable Projection