Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models

First seen · 7/16/2026, 12:28 PMLatest activity · 7/16/2026, 12:28 PM

This paper studies a common repair strategy for reasoning distillation: showing a chain-of-thought generator the gold answer and asking it to produce a derivation reaching that answer. In a controlled experiment, answer-conditioned chains increasingly rationalize backward from the target instead of deriving it. Correctness filtering fails to detect the damage because the final answer is correct. Fine-tuning a strong instruction-tuned reasoning model on its own conditioned chains reduces verifiable-reasoning accuracy, with losses reaching about 27 points on the hardest competition problems. The effect is visible before training, transfers across teacher families, and is attributed to the rationalize-toward instruction rather than answer visibility alone.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/16, 12:28 PMnot independentRepresentative
    Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models