Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Sample-Efficient Learning from Agent Experience

First seen · 7/24/2026, 12:00 PMLatest activity · 7/24/2026, 12:00 PM

The paper introduces Experience Distillation, a method for internalizing an agent’s interaction history into model weights without collecting additional environment interactions. Across 749 curated software-engineering tasks and six text-adventure games, it retained at least 64.8% of the gains achieved by in-context learning, while direct supervised fine-tuning recovered only 3.8%. Against classical reinforcement-learning baselines, in-context learning from trial-and-error experience followed by Experience Distillation matched performance using at least 9.6 times fewer environment samples.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/24, 12:00 PMnot independentRepresentative
    Sample-Efficient Learning from Agent Experience