Codesota · Benchmark · OK-VQAHome/Leaderboards/Multimodal Media/Visual Question Answering/OK-VQA
Unknown

OK-VQA.

14,055 questions requiring outside knowledge to answer. Tests models that must consult external knowledge sources beyond visual content.

Paper ↗Leaderboard ↓Lineage
§ 01 · Leaderboard

Results by metric.

Found a wrong score or missing run?
Use row edits to send a sourced correction into moderation.
Add / edit result ↗Report issue ↗

accuracy

Accuracy is the reported evaluation metric for OK-VQA. Codesota tracks published model scores on this metric so readers can compare state-of-the-art results across sources and model families.

Higher is better

Trust tiers for accuracyverifiedpapervendorcommunityunverified

Muted rows were not state of the art when published — an earlier or same-year result already scored better.

RankModelTrustScoreYearLinksFix
01PaLI-X-55B
Fetched from CodeSOTA API on 2026-04-20
verified66.12026Source ↗Looks wrong?
02PaLI-17B
Fetched from CodeSOTA API on 2026-04-20
verified64.52026Source ↗Looks wrong?
03GPT-4V
Fetched from CodeSOTA API on 2026-04-20
verified64.282026Source ↗Looks wrong?
04Flamingo-80B
Fetched from CodeSOTA API on 2026-04-20
verified57.82026Source ↗Looks wrong?
05BLIP-2 (FlanT5XXL)
Fetched from CodeSOTA API on 2026-04-20
verified44.72026Source ↗Looks wrong?
Lineage

OK-VQA in context.

See full visual question answering lineage →
This benchmark (1)
active2019-06
OK-VQA
Successors (1)
active2022-06
A-OKVQA
Broader knowledge types and better annotation.
§ 04 · Submit a result

Add to the leaderboard.

Submit a Result

Sign in to submit benchmark results for OK-VQA.

Sign in
← Back to Visual Question Answering