Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments

First seen · 7/7/2026, 12:00 PMLatest activity · 7/7/2026, 12:00 PM

EdgeBench studies how deployed agents improve through interaction with real-world environments. Across about 38,000 hours of interaction and 134 ultra-long-horizon tasks, the authors report that aggregate learning performance follows a log-sigmoid scaling law with R² = 0.998. They also observe that learning speed across model generations roughly doubles every three months. Tasks span scientific discovery, software engineering, optimization, professional knowledge work, formal mathematics, and games, with each supporting at least 12 hours of continuous operation. The release includes 51 tasks and the full evaluation framework.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/7, 12:00 PMnot independentRepresentative
    EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments