Codesota · Models2,268 models indexed
Editorial · Models

Models with recorded evidence.

Start with a research area, drill into a vendor, or page through the full index. Vendor aliases and legacy area IDs are grouped here; original model IDs and model links stay unchanged. Only models with at least one benchmark score appear — a model without a recorded score can’t be ranked.

Vendor:Areas overviewSpeakLeash · 263Alibaba · 104Google · 102OpenAI · 86Meta · 68Microsoft · 49Anthropic · 44DeepSeek · 34Mistral · 30mistralai · 19CYFRAGOVPL · 14NVIDIA · 14Zhipu AI · 13internlm · 10xAI · 10ByteDance · 9Baidu · 8ibm-granite · 8PLLuM · 8allenai · 7Amazon · 7MiniMax · 7Mistral AI · 7Remek · 7Shanghai AI Lab · 7utter-project · 7CohereForAI · 6Salesforce · 601-ai · 5Cohere · 5Moonshot AI · 5NousResearch · 5THUML · 5gguf-iq · 4IBM · 4Meituan · 4openchat · 4Stanford · 4THUDM · 4tiiuae · 4UC San Diego · 4VikParuchuri · 4Allen AI · 3BAAI · 3Du et al. · 3ForgeCode · 3Fudan University · 3gguf · 3gguf11bv30 · 3gguf7bv30 · 3IDEA Research · 3Liao et al. · 3Moonshot.AI · 3Nam Tuan Ly / NII · 3OpenDataLab · 3OPI-PG · 3upstage · 3ViCoS Lab Ljubljana · 3Xiaomi · 3Zhao et al. · 3+ 243 smaller vendors (288 models)
§ 01 · Research areas

17 areas, each with a complete model index.

Research area
Computer Vision
896 models · 2,328 results
led by Unknown
Research area
Natural Language Processing
854 models · 7,468 results
led by SpeakLeash
Research area
Agentic AI
164 models · 225 results
led by OpenAI
Research area
Computer Code
152 models · 297 results
led by Anthropic
Research area
Reasoning
151 models · 415 results
led by OpenAI
Research area
Speech
104 models · 532 results
led by Meta
Research area
Multimodal
88 models · 267 results
led by Alibaba
Research area
Medical
50 models · 83 results
led by Research
Research area
Audio
25 models · 32 results
led by Microsoft
Research area
Industrial Inspection
22 models · 27 results
led by Research
Research area
Time Series
21 models · 82 results
led by THUML
Research area
Reinforcement Learning
20 models · 21 results
led by Google
Research area
Graphs
12 models · 12 results
led by Unknown
Research area
Mobile Development
10 models · 40 results
led by Anthropic
Research area
Knowledge Base
9 models · 9 results
led by Meta
Research area
Robots
5 models · 5 results
led by Stanford / Google DeepMind / TRI
Research area
General
3 models · 8 results
led by Unknown

Click an area to see the full paginated list of models scored on its benchmarks. Counts describe recorded evidence, not verified current SOTA. Models are ordered by the number of records; comparisons still require a matched benchmark version and protocol.