Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Scalable Causal Imitation Learning

First seen · 7/19/2026, 07:33 AMLatest activity · 7/19/2026, 07:33 AM

This paper addresses causal imitation learning when expert demonstrations contain unobserved confounders and the expert and imitator have mismatched observations. It introduces Causal Soft Q Imitation Learning (Causal SQIL) and Causal Inverse soft-Q Learning (Causal IQ-Learn), combining causal adjustment with off-policy inverse reinforcement learning objectives. The methods approximate the sequential π-backdoor criterion through a fixed-size sliding window over causally adjusted representations. According to the abstract, evaluations in confounded continuous-control environments show substantially better long-horizon performance than prior Causal BC and Causal GAIL methods, with some results surpassing the expert, while causally unaware baselines fail to learn meaningful behavior.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/19, 07:33 AMnot independentRepresentative
    Scalable Causal Imitation Learning