ParseBench

Overall score versus model parameters. Smaller and higher is better.

ParseBench score versus model sizeLower parameter count and higher overall score are better. Dashed line: Pareto frontier.4045505560657075800.712481632Model parameters (B) · Smaller is betterOverall score · Higher is betterMistral OCR 4.1 — 68.23 points — Size undisclosed — Reported resultsMistral OCR 4.1 · 68.2LightOnOCR-3-4B — 75.1 points — 4.0B — Evaluated outputLightOnOCR-3 4BLightOnOCR-3-0.8B — 74.6 points — 0.8B — Evaluated outputLightOnOCR-3 0.8BInfinity-Parser2-Pro — 74.3 points — 35.1B — Reported resultsInfinity-Parser2-ProLightOnOCR-3-1B — 71.4 points — 1.0B — Evaluated outputLightOnOCR-3-1Bchandra-ocr-2 — 70.1 points — 4.0B — Evaluated outputchandra-ocr-2surya-ocr-2 — 64.8 points — 0.7B — Evaluated outputsurya-ocr-2dots.mocr — 55.8 points — 3.0B — Evaluated outputdots.mocrLightOnOCR-2-1B — 48.0 points — 1.0B — Evaluated outputLightOnOCR-2jina-ocr-v1 — 45.9 points — 3.4B — Evaluated outputjina-ocr-v1
External modelsLightOnOCR-2LightOnOCR-3Pareto frontier
Hover or focus a point for its exact checkpoint and scores.
Method and source

Smaller size and higher score are better. The frontier includes points that no other comparable point matches or improves on in both dimensions, with a strict improvement in at least one. Comparisons use source precision; displayed scores are rounded to one decimal. Parameter counts are those in the tables, including nominal backbone sizes where specified.

Overall uses five categories for ParseBench and includes headers/footers for olmOCR. All points match the selected table rows; LightOnOCR-3 uses the candidate-1 evaluations for 0.8B, 1B, and 4B. Postprocessing applies only to olmOCR; ParseBench and fr-bench-pdf2md outputs are evaluated as submitted. The same candidate checkpoints are used across all three benchmarks. Missing scores or sizes are omitted. Models with undisclosed parameter counts (such as MistralOCR4.1) cannot be placed on a size axis. * LightOnOCR-2 uses its published olmOCR score excluding headers/footers; it is shown but excluded from the frontier. Mistral OCR 4 is a horizontal score reference because its size is undisclosed; it is excluded from the size frontier. Its olmOCR score is the user-provided reported 85.20; ParseBench and fr-bench-pdf2md retain their existing selected scores.

Leaderboard source revision: 3576c5240602588f596f8790a3a7161a524e59ad. Values are pinned to this snapshot.

Show scores and output details
Plotted scores
ModelSize (B)OverallOutput / inputsFrontier
LightOnOCR-3-4B4.075.1Evaluated output✓
LightOnOCR-3-0.8B0.874.6Evaluated output✓
Infinity-Parser2-Pro35.174.3Reported results—
LightOnOCR-3-1B1.071.4Evaluated output—
chandra-ocr-24.070.1Evaluated output—
Mistral OCR 4.1Undisclosed68.2Reported resultsSize undisclosed
surya-ocr-20.764.8Evaluated output✓
dots.mocr3.055.8Evaluated output—
LightOnOCR-2-1B1.048.0Evaluated output—
jina-ocr-v13.445.9Evaluated output—