Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember

First seen · 7/31/2026, 10:32 PMLatest activity · 7/31/2026, 10:32 PM

The paper introduces SESA, a Self-Evolving Skill-Augmented Agent in which a challenger poses problems, a separately parameterized solver retrieves procedural skills, and informative failures are distilled back into an evolving memory bank. Because memory changes solver trajectories, policy learning, and the challenger’s future problem distribution, task generation and skill memory co-evolve. Across seven open-domain and multi-hop QA benchmarks, SESA improves average accuracy over SSP by 1.2–3.2 points and exceeds SkillRL by 0.9 points under a unified protocol. On Qwen3, memory-free SESA-Off retains a 1.8–2.2-point gain over SSP, while inference-time retrieval adds 0.5–1.0 points.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/31, 10:32 PMnot independentRepresentative
    Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember