Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Can AI Agents Simulate A/B Test Outcomes? A Validation Framework for Agentic Experimentation

First seen · 8/3/2026, 10:58 PMLatest activity · 8/3/2026, 10:58 PM

This paper formalizes AI-agent-based experiment simulation as a Simulated Randomized Controlled Trial (S-RCT), with an error decomposition separating agent approximation error from subsampling error. Evaluated on 67 historical marketing A/B tests, an off-the-shelf foundation-model baseline achieved a sign overlap of 0.70 but systematically overstated effect sizes. A two-phase pre-period calibration protocol reduced squared prediction error, after removing irreducible measurement noise, by approximately 77x. A within-subject design reduced standard errors by approximately 2.4x. The authors present the framework as agent-agnostic and discuss where simulated signals may assist experiment planning.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv8/3, 10:58 PMnot independentRepresentative
    Can AI Agents Simulate A/B Test Outcomes? A Validation Framework for Agentic Experimentation