BENCHMARKS Chandra leads open-source OCR. The Datalab API ships a tuned variant. ← All benchmarks
TABLES · OLMOCR-BENCH

Tables — 90.7% on olmOCR-bench.

olmOCR-bench unit-tests OCR output against known-correct table cells in real-world documents — financial filings, scientific papers, regulatory reports. We focus on the structurally tricky cases competitors flub.

SCOREBOARD · olmOCR-bench

Ranked scoreboard.

Models ranked highest-to-lowest. Datalab variants in accent; competitors and prior generations in muted ink.

Rank Model Score vs scale
01 Datalab API 90.7%
02 Chandra 2 OSS 89.9%
03 Chandra 1 prior generation 88.0%
Dataset · olmOCR-bench Last run · 2026-03-18

+1.9 vs prior generation

WHY THIS NUMBER IS CREDIBLE

olmOCR-bench unit-tests OCR output against known-correct elements in a public dataset. Model versions pinned per run; anyone can reproduce from the olmOCR-bench HuggingFace dataset.

START

Nested headers, merged cells, colspans.

Run your hardest table on Chandra — free tier, no credit card.