Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation

First seen · 7/31/2026, 11:35 PMLatest activity · 7/31/2026, 11:35 PM

The paper introduces the Authorship-Rewriting Benchmark (ARB), a matched dataset built from 1,800 human texts drawn from XSum, WritingPrompts, and OpenWebText. Each source produces four variants: human-written text, direct LLM generation, human-to-LLM rewriting, and same-generator rewriting of LLM text. At a strict 1% false-positive rate, FastDetectGPT and Binoculars-falcon-7b detected 91.2% and 93.5% of direct LLM text, but only 30.8% and 15.1% of human text rewritten by an LLM. Detection of LLM-originated text remained substantially stronger after rewriting.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/31, 11:35 PMnot independentRepresentative
    ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation