Multi-scene text reading, key information extraction, multilingual text, and document parsing benchmark.
Multi Scene F1 is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.
Higher is better
Muted rows were not state of the art when published — an earlier or same-year result already scored better.
| Rank | Model | Trust | Score | Year | Links | Fix |
|---|---|---|---|---|---|---|
| 01 | gemini-15-pro | unverified | 83.25 | 2026 | N/A | Looks wrong? |
| 02 | qwen2-vl-72b | unverified | 77.95 | 2026 | N/A | Looks wrong? |
| 03 | internvl2-76b | unverified | 76.92 | 2026 | N/A | Looks wrong? |
| 04 | gpt-4o | unverified | 76.4 | 2026 | N/A | Looks wrong? |
| 05 | claude-35-sonnet | unverified | 72.87 | 2026 | N/A | Looks wrong? |
Multilingual F1 is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.
Higher is better
Muted rows were not state of the art when published — an earlier or same-year result already scored better.
| Rank | Model | Trust | Score | Year | Links | Fix |
|---|---|---|---|---|---|---|
| 01 | gemini-15-pro | unverified | 78.97 | 2026 | N/A | Looks wrong? |
| 02 | gpt-4o | unverified | 73.44 | 2026 | N/A | Looks wrong? |
Kie F1 is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.
Higher is better
Muted rows were not state of the art when published — an earlier or same-year result already scored better.
| Rank | Model | Trust | Score | Year | Links | Fix |
|---|---|---|---|---|---|---|
| 01 | qwen2-vl-72b | unverified | 71.76 | 2026 | N/A | Looks wrong? |
| 02 | gemini-15-pro | unverified | 67.28 | 2026 | N/A | Looks wrong? |
| 03 | claude-35-sonnet | unverified | 64.58 | 2026 | N/A | Looks wrong? |
| 04 | gpt-4o | unverified | 63.45 | 2026 | N/A | Looks wrong? |
Document Parsing is the reported evaluation metric for CC-OCR. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.
Higher is better
Muted rows were not state of the art when published — an earlier or same-year result already scored better.
| Rank | Model | Trust | Score | Year | Links | Fix |
|---|---|---|---|---|---|---|
| 01 | gemini-15-pro | unverified | 62.37 | 2026 | N/A | Looks wrong? |