Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

One More Turn, Less Regret: A Regret-Based Multi-Turn Benchmark for LLMs' Clarification Policies

First seen · 7/23/2026, 06:22 PMLatest activity · 7/23/2026, 06:22 PM

The paper introduces RegretBench, a benchmark for evaluating clarification as a sequential policy rather than judging individual questions in isolation. It uses hidden user intents, free-form interaction, semantic-state tracking, and a regret objective that measures value lost relative to a reference clarification policy. Experiments cover open-domain question answering and product recommendation. The reported results suggest that final task success alone misses important differences: models with similar accuracy can vary in interaction efficiency, robustness to user behavior, ineffective questioning, and stopping decisions. Useful clarification depends on asking the right question at the right time and stopping when the intended meaning is sufficiently resolved.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/23, 06:22 PMnot independentRepresentative
    One More Turn, Less Regret: A Regret-Based Multi-Turn Benchmark for LLMs' Clarification Policies