Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

DADiff: Diffusion-Driven Cross-Domain Policy Adaptation for Reinforcement Learning

First seen · 7/18/2026, 12:20 AMLatest activity · 7/18/2026, 12:20 AM

DADiff addresses online dynamics adaptation in reinforcement learning, where a policy is trained with abundant source-domain data but can collect only limited target-domain interactions. The method uses a diffusion-based generative process for next-state prediction and estimates dynamics mismatch from discrepancies between source- and target-domain generative trajectories. It offers two adaptation variants: reward modification and data selection. The paper also derives a theoretical bound relating policy performance differences across domains to generative trajectory deviation, and reports experiments across multiple types of domain shifts.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/18, 12:20 AMnot independentRepresentative
    DADiff: Diffusion-Driven Cross-Domain Policy Adaptation for Reinforcement Learning