Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

OvisOCR2 Technical Report: A 0.8B End-to-End Document Parser Tops OmniDocBench

First seen · 7/16/2026, 12:00 PMLatest activity · 7/16/2026, 12:00 PM

OvisOCR2 is a 0.8B end-to-end document parsing model that converts page images into Markdown in natural reading order, covering text, formulas, tables, and visual regions. Its training pipeline combines filtered real-document annotations, HTML-derived synthetic pages, supervised fine-tuning, reinforcement learning on a 4B branch, on-policy distillation, and model fusion. The report states that OvisOCR2 reaches 96.58 on OmniDocBench v1.6 and 75.06 Avg3 on PureDocBench, with a public Hugging Face release.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/15, 05:32 PMnot independent
    OvisOCR2 Technical Report
  2. AggregatorHuggingFace Daily Papers7/16, 12:00 PMnot independentRepresentative
    OvisOCR2 Technical Report: A 0.8B End-to-End Document Parser Tops OmniDocBench