Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Necessary or Sufficient? Evaluating LLM Explanations with Behavioural Evidence

First seen · 9/5/2026, 01:37 AMLatest activity · 9/5/2026, 01:37 AM

When LLMs make decisions within agent workflows, they often cite the primary factors behind their judgements. Evaluating eight models across the Claude, GPT, and Gemini families through controlled black-box interventions, this study tests whether these cited explanations meet behavioural standards of necessity and sufficiency. In an advisor-recommendation benchmark, uncited factors held more measurable sway than the lowest-ranked cited factor in over 57% of cases. The findings indicate that while self-reported explanations carry partial signal, they cannot be reliably trusted as true causal accounts for safety auditing.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv9/5, 01:37 AMnot independentRepresentative
    Necessary or Sufficient? Evaluating LLM Explanations with Behavioural Evidence