Codesota · Benchmark · CC-OCRHome/Leaderboards/CC-OCR
South China University of Technology

CC-OCR.

Multi-scene text reading, key information extraction, multilingual text, and document parsing benchmark.

Paper ↗Leaderboard ↓
§ 01 · Leaderboard

Results by metric.

Found a wrong score or missing run?
Use row edits to send a sourced correction into moderation.
Add / edit result ↗Report issue ↗

Multi-Scene F1

Multi Scene F1 is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Multi-Scene F1verifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01gemini-15-pro
Non-API entry from src
unverified83.252026N/ALooks wrong?
02qwen2-vl-72b
Non-API entry from src
unverified77.952026N/ALooks wrong?
03internvl2-76b
Non-API entry from src
unverified76.922026N/ALooks wrong?
04gpt-4o
Non-API entry from src
unverified76.42026N/ALooks wrong?
05claude-35-sonnet
Non-API entry from src
unverified72.872026N/ALooks wrong?

Multilingual F1

Multilingual F1 is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Multilingual F1verifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01gemini-15-pro
Non-API entry from src
unverified78.972026N/ALooks wrong?
02gpt-4o
Non-API entry from src
unverified73.442026N/ALooks wrong?

KIE F1

Kie F1 is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for KIE F1verifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01qwen2-vl-72b
Non-API entry from src
unverified71.762026N/ALooks wrong?
02gemini-15-pro
Non-API entry from src
unverified67.282026N/ALooks wrong?
03claude-35-sonnet
Non-API entry from src
unverified64.582026N/ALooks wrong?
04gpt-4o
Non-API entry from src
unverified63.452026N/ALooks wrong?

Document Parsing

Document Parsing is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for Document Parsingverifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01gemini-15-pro
Non-API entry from src
unverified62.372026N/ALooks wrong?
§ 04 · Submit a result

Add to the leaderboard.

Submit a Result

Sign in to submit benchmark results for CC-OCR.

Sign in
← Back to Leaderboards