Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Reasoning Consistency Scanning: A Framework for Auditing Chain-of-Thought Validity in AI Safety Evaluations

First seen · 7/8/2026, 06:13 PMLatest activity · 7/8/2026, 06:13 PM

This paper separates chain-of-thought faithfulness from reasoning consistency. Faithfulness requires controlled interventions to test whether stated reasoning reflects the process that generated an answer, while consistency can be assessed from a transcript alone. The authors define a six-subtype taxonomy of inconsistency, create a validated benchmark of 60 manually adapted transcripts from InstrumentalEval outputs, and implement a scanner in InspectScout. Experiments cover four generator models and three evaluations from inspect_evals, reporting that inconsistency is present, detectable, and varies systematically by model and task type.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/8, 06:13 PMnot independentRepresentative
    Reasoning Consistency Scanning: A Framework for Auditing Chain-of-Thought Validity in AI Safety Evaluations