Codesota · Benchmark · OCRBench v2Home/Leaderboards/OCRBench v2
South China University of Technology

OCRBench v2.

Tests 8 core OCR capabilities across 23 tasks. Evaluates LMMs on text recognition, referring, extraction.

Paper ↗Leaderboard ↓Lineage
§ 01 · Leaderboard

Results by metric.

Found a wrong score or missing run?
Use row edits to send a sourced correction into moderation.
Add / edit result ↗Report issue ↗

Overall (Chinese)

Overall Zh Private is the reported evaluation metric for OCRBench v2. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Overall (Chinese)verifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01Qwen2.5-VL-72B
Fetched from CodeSOTA API on 2026-04-20
vendor63.72026Source ↗Looks wrong?
02gemini-25-pro
Fetched from CodeSOTA API on 2026-04-20
vendor62.22026Source ↗Looks wrong?
03Qianfan-OCR
Fetched from CodeSOTA API on 2026-04-20
vendor60.772026Source ↗Looks wrong?
04intern-s1-pro
Mapped from PWC OCRBench v2 Chinese Score.; Reported in the Intern-S1-Pro paper and Hugging Face model card performance table as OCRBench V2 (ENG / CHN). OCRBench V2 is evaluated with the non-thinking configuration; scores are English 60.1 and Chinese 60.6.; PWC evaluation id 5083; paper: Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale
verified60.62026Source ↗Looks wrong?
05minicpm-v-4.5-8b
Fetched from CodeSOTA API on 2026-04-20
vendor58.82026Source ↗Looks wrong?
06ovis2-5-9b
Mapped from PWC OCRBench v2 Chinese Score.; Table 6, OCR & chart; OCRBench v2 Chinese split. Source/provenance: Ovis2.5 Technical Report; source arXiv paper https://arxiv.org/abs/2508.11737; official HF model URL https://huggingface.co/AIDC-AI/Ovis2.5-9B.; PWC evaluation id 5587; paper: Ovis2.5 Technical Report
verified582026Source ↗Looks wrong?
07sail-vl2-8b
Fetched from CodeSOTA API on 2026-04-20
vendor57.62026Source ↗Looks wrong?
08claude-3.5-sonnet
Fetched from CodeSOTA API on 2026-04-20
vendor48.42026Source ↗Looks wrong?
09InternVL2.5-78B
Fetched from CodeSOTA API on 2026-04-20
vendor46.22026Source ↗Looks wrong?
10Qwen2-VL-72B
Fetched from CodeSOTA API on 2026-04-20
vendor46.12026Source ↗Looks wrong?
11gpt-4o-2024
Fetched from CodeSOTA API on 2026-04-20
vendor45.72026Source ↗Looks wrong?

Overall (English)

Overall En Private is the reported evaluation metric for OCRBench v2. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Overall (English)verifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01ovis2-5-9b
Mapped from PWC OCRBench v2 English Score.; Table 6, OCR & chart; OCRBench v2 English split. Source/provenance: Ovis2.5 Technical Report; source arXiv paper https://arxiv.org/abs/2508.11737; official HF model URL https://huggingface.co/AIDC-AI/Ovis2.5-9B.; PWC evaluation id 5586; paper: Ovis2.5 Technical Report
verified63.42026Source ↗Looks wrong?
02seed-1.6-vision
Fetched from CodeSOTA API on 2026-04-20
vendor62.22026Source ↗Looks wrong?
03Qwen2.5-VL-72B
Fetched from CodeSOTA API on 2026-04-20
vendor61.52026Source ↗Looks wrong?
04qwen3-omni-30b
Fetched from CodeSOTA API on 2026-04-20
vendor61.32026Source ↗Looks wrong?
05nemotron-nano-v2-vl
Fetched from CodeSOTA API on 2026-04-20
vendor61.22026Source ↗Looks wrong?
06intern-s1-pro
Mapped from PWC OCRBench v2 English Score.; Reported in the Intern-S1-Pro paper and Hugging Face model card performance table as OCRBench V2 (ENG / CHN). OCRBench V2 is evaluated with the non-thinking configuration; scores are English 60.1 and Chinese 60.6.; PWC evaluation id 5083; paper: Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale
verified60.12026Source ↗Looks wrong?
07gemini-25-pro
Fetched from CodeSOTA API on 2026-04-20
vendor59.32026Source ↗Looks wrong?
08llama-3.1-nemotron-nano-vl-8b
Fetched from CodeSOTA API on 2026-04-20
vendor56.42026Source ↗Looks wrong?
09Qianfan-OCR
Fetched from CodeSOTA API on 2026-04-20
vendor562026Source ↗Looks wrong?
10gpt-4o
Fetched from CodeSOTA API on 2026-04-20
vendor55.52026Source ↗Looks wrong?
11ovis2.5-8b
Fetched from CodeSOTA API on 2026-04-20
vendor54.12026Source ↗Looks wrong?
12gemini-1.5-pro
Fetched from CodeSOTA API on 2026-04-20
vendor51.62026Source ↗Looks wrong?
13sail-vl2-8b
Fetched from CodeSOTA API on 2026-04-20
vendor49.32026Source ↗Looks wrong?
14minicpm-v-4.5-8b
Fetched from CodeSOTA API on 2026-04-20
vendor48.42026Source ↗Looks wrong?
15Qwen2-VL-72B
Fetched from CodeSOTA API on 2026-04-20
vendor47.82026Source ↗Looks wrong?
16gpt-4o-2024
Fetched from CodeSOTA API on 2026-04-20
vendor47.62026Source ↗Looks wrong?
17claude-3.5-sonnet
Fetched from CodeSOTA API on 2026-04-20
vendor47.52026Source ↗Looks wrong?
18internvl3.5-14b
Fetched from CodeSOTA API on 2026-04-20
vendor47.12026Source ↗Looks wrong?
19step-1v
Fetched from CodeSOTA API on 2026-04-20
vendor46.82026Source ↗Looks wrong?
20InternVL2.5-78B
Fetched from CodeSOTA API on 2026-04-20
vendor452026Source ↗Looks wrong?
21grok4
Fetched from CodeSOTA API on 2026-04-20
vendor452026Source ↗Looks wrong?
22gpt-4o-mini
Fetched from CodeSOTA API on 2026-04-20
vendor44.12026Source ↗Looks wrong?
23claude-sonnet-4
Fetched from CodeSOTA API on 2026-04-20
vendor42.42026Source ↗Looks wrong?
24qwen2.5-vl-7b
Fetched from CodeSOTA API on 2026-04-20
vendor41.82026Source ↗Looks wrong?
25deepseek-vl2-small
Fetched from CodeSOTA API on 2026-04-20
vendor412026Source ↗Looks wrong?
26pixtral-12b
Fetched from CodeSOTA API on 2026-04-20
vendor38.42026Source ↗Looks wrong?
27phi-4-multimodal
Fetched from CodeSOTA API on 2026-04-20
vendor38.12026Source ↗Looks wrong?
28glm-4v-9b
Fetched from CodeSOTA API on 2026-04-20
vendor37.12026Source ↗Looks wrong?
29molmo-7b
Fetched from CodeSOTA API on 2026-04-20
vendor33.92026Source ↗Looks wrong?
30llava-ov-7b
Fetched from CodeSOTA API on 2026-04-20
vendor33.72026Source ↗Looks wrong?
31idefics3-8b
Fetched from CodeSOTA API on 2026-04-20
vendor262026Source ↗Looks wrong?
32mistral-ocr-2512
Fetched from CodeSOTA API on 2026-04-20
verified25.22026Source ↗Looks wrong?
33docowl2
Fetched from CodeSOTA API on 2026-04-20
vendor23.42026Source ↗Looks wrong?

Overall Zh Public

Overall Zh Public is the reported evaluation metric for OCRBench v2. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Overall Zh Publicverifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01InternVL3-14B
Fetched from CodeSOTA API on 2026-04-20
vendor55.72026Source ↗Looks wrong?
02Qwen2.5-VL-7B
Fetched from CodeSOTA API on 2026-04-20
vendor55.62026Source ↗Looks wrong?
03Ovis2-8B
Fetched from CodeSOTA API on 2026-04-20
vendor49.22026Source ↗Looks wrong?
04Gemini 1.5 Pro
Fetched from CodeSOTA API on 2026-04-20
vendor43.12026Source ↗Looks wrong?
05DeepSeek-VL2-Small
Fetched from CodeSOTA API on 2026-04-20
vendor42.72026Source ↗Looks wrong?
06Step-1V
Fetched from CodeSOTA API on 2026-04-20
vendor42.62026Source ↗Looks wrong?
07MiniCPM-o-2.6
Fetched from CodeSOTA API on 2026-04-20
vendor41.12026Source ↗Looks wrong?
08Claude 3.5 Sonnet
Fetched from CodeSOTA API on 2026-04-20
vendor39.62026Source ↗Looks wrong?
09GLM-4V-9B
Fetched from CodeSOTA API on 2026-04-20
vendor36.62026Source ↗Looks wrong?
10GPT-4o
Fetched from CodeSOTA API on 2026-04-20
vendor32.22026Source ↗Looks wrong?
11LLaVA-OneVision-7B
Fetched from CodeSOTA API on 2026-04-20
vendor17.82026Source ↗Looks wrong?
12TextMonkey
Fetched from CodeSOTA API on 2026-04-20
vendor15.82026Source ↗Looks wrong?
13Pixtral-12B
Fetched from CodeSOTA API on 2026-04-20
vendor14.62026Source ↗Looks wrong?
14Monkey
Fetched from CodeSOTA API on 2026-04-20
vendor13.12026Source ↗Looks wrong?
15Molmo-7B
Fetched from CodeSOTA API on 2026-04-20
vendor12.82026Source ↗Looks wrong?
16Cambrian-1-8B
Fetched from CodeSOTA API on 2026-04-20
vendor9.902026Source ↗Looks wrong?
17LLaVA-NeXT-8B
Fetched from CodeSOTA API on 2026-04-20
vendor9.102026Source ↗Looks wrong?

Overall En Public

Overall En Public is the reported evaluation metric for OCRBench v2. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Overall En Publicverifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01InternVL3-14B
Fetched from CodeSOTA API on 2026-04-20
vendor52.62026Source ↗Looks wrong?
02Gemini 1.5 Pro
Fetched from CodeSOTA API on 2026-04-20
vendor51.92026Source ↗Looks wrong?
03Ovis2-8B
Fetched from CodeSOTA API on 2026-04-20
vendor47.72026Source ↗Looks wrong?
04Qwen2.5-VL-7B
Fetched from CodeSOTA API on 2026-04-20
vendor46.72026Source ↗Looks wrong?
05Step-1V
Fetched from CodeSOTA API on 2026-04-20
vendor46.72026Source ↗Looks wrong?
06GPT-4o
Fetched from CodeSOTA API on 2026-04-20
vendor46.52026Source ↗Looks wrong?
07Claude 3.5 Sonnet
Fetched from CodeSOTA API on 2026-04-20
vendor45.22026Source ↗Looks wrong?
08MiniCPM-o-2.6
Fetched from CodeSOTA API on 2026-04-20
vendor45.12026Source ↗Looks wrong?
09DeepSeek-VL2-Small
Fetched from CodeSOTA API on 2026-04-20
vendor43.32026Source ↗Looks wrong?
10GLM-4V-9B
Fetched from CodeSOTA API on 2026-04-20
vendor42.62026Source ↗Looks wrong?
11Pixtral-12B
Fetched from CodeSOTA API on 2026-04-20
vendor40.32026Source ↗Looks wrong?
12LLaVA-OneVision-7B
Fetched from CodeSOTA API on 2026-04-20
vendor36.42026Source ↗Looks wrong?
13Cambrian-1-8B
Fetched from CodeSOTA API on 2026-04-20
vendor34.72026Source ↗Looks wrong?
14Molmo-7B
Fetched from CodeSOTA API on 2026-04-20
vendor34.52026Source ↗Looks wrong?
15LLaVA-NeXT-8B
Fetched from CodeSOTA API on 2026-04-20
vendor31.52026Source ↗Looks wrong?
16TextMonkey
Fetched from CodeSOTA API on 2026-04-20
vendor23.92026Source ↗Looks wrong?
17Monkey
Fetched from CodeSOTA API on 2026-04-20
vendor23.12026Source ↗Looks wrong?
Lineage

OCRBench v2 in context.

See full ocr benchmarks lineage →
Predecessors (1)
superseded2023-05
OCRBench
10× more items, human-verified, EN+ZH parity, four public/private splits to combat contamination. Original v1 saturated within 18 months; v2 reopened the gap.
This benchmark (1)
active2024-12
OCRBench v2
§ 04 · Submit a result

Add to the leaderboard.

Submit a Result

Sign in to submit benchmark results for OCRBench v2.

Sign in
← Back to Leaderboards