Codesota · Benchmark · FUNSDHome/Leaderboards/FUNSD
Unknown

FUNSD.

199 fully annotated forms. Tests semantic entity labeling and linking.

Paper ↗Leaderboard ↓Lineage
§ 01 · Leaderboard

Results by metric.

Found a wrong score or missing run?
Use row edits to send a sourced correction into moderation.
Add / edit result ↗Report issue ↗

f1

F1 is the reported evaluation metric for FUNSD. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for f1verifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01LayoutLMv3-large
Fetched from CodeSOTA API on 2026-04-20
verified92.082026Source ↗Looks wrong?
02UDOP
Fetched from CodeSOTA API on 2026-04-20
verified91.622026Source ↗Looks wrong?
03LayoutLMv3-base
Fetched from CodeSOTA API on 2026-04-20
verified90.292026Source ↗Looks wrong?
04DocFormerv2-large
Fetched from CodeSOTA API on 2026-04-20
verified88.892026Source ↗Looks wrong?
05LiLT[EN-R2]-base
Fetched from CodeSOTA API on 2026-04-20
verified88.412026Source ↗Looks wrong?
06DocFormerv2-base
Fetched from CodeSOTA API on 2026-04-20
verified88.372026Source ↗Looks wrong?
07StructuralLM
Fetched from CodeSOTA API on 2026-04-20
verified85.142026Source ↗Looks wrong?
08FormNet
Fetched from CodeSOTA API on 2026-04-20
verified84.692026Source ↗Looks wrong?
09BROS-large
Fetched from CodeSOTA API on 2026-04-20
verified84.522026Source ↗Looks wrong?
10LayoutLMv2-large
Fetched from CodeSOTA API on 2026-04-20
verified84.22026Source ↗Looks wrong?
11LayoutLMv2-base
Fetched from CodeSOTA API on 2026-04-20
verified82.762026Source ↗Looks wrong?
12LayoutLMv1-base
Fetched from CodeSOTA API on 2026-04-20
verified79.272026Source ↗Looks wrong?
13LayoutLMv1-large
Fetched from CodeSOTA API on 2026-04-20
verified77.892026Source ↗Looks wrong?
Lineage

FUNSD in context.

See full ocr benchmarks lineage →
This benchmark (1)
saturated2019-05
FUNSD
Successors (1)
superseded2023-05
OCRBench
Once VLMs could read at all, evaluation needed to span more than forms. OCRBench bundled scene text, document VQA, KIE and handwritten math into one composite — the first VLM-era OCR benchmark.
§ 04 · Submit a result

Add to the leaderboard.

Submit a Result

Sign in to submit benchmark results for FUNSD.

Sign in
← Back to Leaderboards