Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

First seen · 7/2/2026, 12:00 PMLatest activity · 7/2/2026, 12:00 PM

This paper presents a systematic evaluation of data referencing errors (DREs) in large language models performing table tasks. Across models ranging from 1.7B to 20B parameters, the authors report that every tested model made errors such as citing incorrect values or omitting relevant entries, even when it understood the table structure. Treating data referencing as a critic improved answer accuracy by up to 12.0% through critic-based filtering and rejection sampling. A lightweight 4B-parameter critic model achieved an average F1 score of 78.2% for detecting in-distribution and out-of-distribution DREs.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorHuggingFace Daily Papers7/2, 12:00 PMnot independentRepresentative
    When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors