benchgap
Vision & documents

CC-OCR leaderboard

As of 2026-10-10, the highest measured score on CC-OCR is 82.0% by Qwen3.5 397B. 15 more models have estimated scores, calibrated from the benchmarks they were measured on.

Measured scores: benchlm.ai.

#ModelScoreSource
1Command A+88.6%estimated ± 4.5 pp, low confidence
2Qwen3.5-35B-A3B82.7%estimated ± 4.5 pp, low confidence
3MiniMax M382.5%estimated ± 3.7 pp, medium confidence
4Qwen3.7 Plus82.2%estimated ± 3.7 pp, medium confidence
5Qwen3.5 397B82.0%measured
6Qwen3.6-35B-A3B81.9%measured
7Qwen3.5-27B81.7%estimated ± 4.5 pp, low confidence
8Qwen3.8-27B81.7%estimated ± 3.7 pp, medium confidence
9Qwen3.6-27B81.2%measured
10Nemotron 3 Nano Omni 30B A3B81.0%estimated ± 0.7 pp, medium confidence
11LFM2.5-VL-3B80.1%estimated ± 0.7 pp, medium confidence
12Qwen3.5-122B-A10B79.9%estimated ± 4.5 pp, low confidence
13Kimi K2.579.7%measured
14Qwen3.8 Max79.6%measured
15Gemini 3 Pro79.0%measured
16ZAYA1-VL-8B78.9%estimated ± 0.7 pp, medium confidence
17Interfaze Beta78.1%estimated ± 0.7 pp, low confidence
18Qwen3.6 Plus77.4%estimated ± 4.5 pp, low confidence
19Claude Opus 4.576.1%estimated ± 3.7 pp, medium confidence
20GPT-5.270.3%measured
21LFM2.5-VL-450M66.2%estimated ± 4.4 pp, low confidence
22Muse Glimmer 30B58.9%estimated ± 3.7 pp, low confidence

All leaderboards

Agentic · terminal

Agentic · tools

Coding

Math

Knowledge & reasoning

Instruction following

Multilingual

Vision & documents

Long context

General