Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Step-Level Preference Learning for Generative Agents in Social Simulations

First seen · 7/16/2026, 09:58 AMLatest activity · 7/16/2026, 09:58 AM

The paper introduces an interactive simulation interface for collecting human preferences over intermediate decisions in generative-agent trajectories, including planning, memory retrieval, reflection, and action selection. It releases a dataset containing 57K fine-grained annotations and applies supervised fine-tuning and direct preference optimization to open-weight language models. According to the abstract, both methods consistently improve simulation fidelity, coordination, interaction quality, and socially effective behavior. The work frames step-level preference supervision as a training signal for improving local decisions while also influencing long-horizon agent behavior.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/16, 09:58 AMnot independentRepresentative
    Step-Level Preference Learning for Generative Agents in Social Simulations