Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

ORCAID: Oblique Rule-Based Continuous-Action Interpretation for Deep RL Policies

First seen · 7/8/2026, 06:16 PMLatest activity · 7/8/2026, 06:16 PM

ORCAID extracts interpretable, rule-based policies from deep reinforcement learning agents in mixed continuous-discrete environments with continuous actions. It trains oblique decision trees that partition the state space with hyperplanes and fit local linear models in the leaves. Its three-stage split search combines efficient random initialization, local refinement, and backward elimination, followed by adjacent-leaf merging to produce concise rules. The authors report strong retained performance across multiple RL environments and claim that the distilled policy can sometimes improve the original deep RL policy.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/8, 06:16 PMnot independentRepresentative
    ORCAID: Oblique Rule-Based Continuous-Action Interpretation for Deep RL Policies