Codesota · Models2,268 models indexed · 896 match filter
Editorial · Models
Models with recorded evidence.
Start with a research area, drill into a vendor, or page through the full index. Vendor aliases and legacy area IDs are grouped here; original model IDs and model links stay unchanged. Only models with at least one benchmark score appear — a model without a recorded score can’t be ranked.
Vendor:Areas overviewSpeakLeash · 263Alibaba · 104Google · 102OpenAI · 86Meta · 68Microsoft · 49Anthropic · 44DeepSeek · 34Mistral · 30mistralai · 19CYFRAGOVPL · 14NVIDIA · 14Zhipu AI · 13internlm · 10xAI · 10ByteDance · 9Baidu · 8ibm-granite · 8PLLuM · 8allenai · 7Amazon · 7MiniMax · 7Mistral AI · 7Remek · 7Shanghai AI Lab · 7utter-project · 7CohereForAI · 6Salesforce · 601-ai · 5Cohere · 5Moonshot AI · 5NousResearch · 5THUML · 5gguf-iq · 4IBM · 4Meituan · 4openchat · 4Stanford · 4THUDM · 4tiiuae · 4UC San Diego · 4VikParuchuri · 4Allen AI · 3BAAI · 3Du et al. · 3ForgeCode · 3Fudan University · 3gguf · 3gguf11bv30 · 3gguf7bv30 · 3IDEA Research · 3Liao et al. · 3Moonshot.AI · 3Nam Tuan Ly / NII · 3OpenDataLab · 3OPI-PG · 3upstage · 3ViCoS Lab Ljubljana · 3Xiaomi · 3Zhao et al. · 3+ 243 smaller vendors (288 models)
§ 01 · Computer Vision models
896 models in Computer Vision · page 11 of 18.
| # | Model | Vendor | Parameters | Architecture | Benchmarks | Results |
|---|---|---|---|---|---|---|
| 501 | BERT [BERT] | Unknown | Unknown | Unknown | 1 | 1 |
| 502 | Binder | Unknown | Unknown | Unknown | 1 | 1 |
| 503 | BioGPT-Large | — | — | — | 1 | 1 |
| 504 | BioRex+Directionality | Unknown | Unknown | Unknown | 1 | 1 |
| 505 | Bluche | Unknown | Unknown | Unknown | 1 | 1 |
| 506 | BM25-HierSumm (query: step + method + article titles) | Unknown | Unknown | Unknown | 1 | 1 |
| 507 | BM25-HierSumm (query: step + method titles) | Unknown | Unknown | Unknown | 1 | 1 |
| 508 | BM25-HierSumm (query: step title) | Unknown | Unknown | Unknown | 1 | 1 |
| 509 | cascadetabnet | Unknown | Unknown | Unknown | 1 | 1 |
| 510 | CCD-ViT-Tiny | Unknown | Unknown | Unknown | 1 | 1 |
| 511 | CDeC-Net | Unknown | Unknown | Unknown | 1 | 1 |
| 512 | CDeCNet | Unknown | Unknown | Unknown | 1 | 1 |
| 513 | CDistNet | Research | Unknown | Content & Spatial Distribution Network for scene text recognition | 1 | 1 |
| 514 | CES (query: method + article + steps titles) | Unknown | Unknown | Unknown | 1 | 1 |
| 515 | CES (query: method + article titles) | Unknown | Unknown | Unknown | 1 | 1 |
| 516 | CES (query: method title) | Unknown | Unknown | Unknown | 1 | 1 |
| 517 | CES (query: step + method + article titles) | Unknown | Unknown | Unknown | 1 | 1 |
| 518 | CES (query: step + method titles) | Unknown | Unknown | Unknown | 1 | 1 |
| 519 | CES (query: step title) | Unknown | Unknown | Unknown | 1 | 1 |
| 520 | Chain-of-Table | Unknown | Unknown | Unknown | 1 | 1 |
| 521 | Chandra | — | — | — | 1 | 1 |
| 522 | Chandra 2 | — | — | — | 1 | 1 |
| 523 | CLIP | — | — | — | 1 | 1 |
| 524 | CLIP4STR | Research | Unknown | CLIP-based Scene Text Recognition | 1 | 1 |
| 525 | CLIP4STR-B (DataComp-1B) | Unknown | Unknown | Unknown | 1 | 1 |
| 526 | CLIP4STR-H (DFN-5B) | Zhao et al. | Unknown | CLIP ViT-H/14 visual branch + cross-modal branch, pre-trained on DFN-5B | 1 | 1 |
| 527 | CLIP4STR-L (RBU 6.5M) | Zhao et al. | Unknown | CLIP ViT-L/14 visual branch + cross-modal branch, trained on RBU 6.5M real data | 1 | 1 |
| 528 | CNN | Unknown | Unknown | Unknown | 1 | 1 |
| 529 | CNN + BLSTM | Unknown | Unknown | Unknown | 1 | 1 |
| 530 | coatnet_2_rw_224.sw_in12k_ft_in1k | — | CoAtNet-2 RW, IN12K -> IN1K fine-tune | 1 | 1 | |
| 531 | CoCa (finetuned) | 2.1B | Contrastive Captioner | 1 | 1 | |
| 532 | CoCa (ViT-G/14) | 2.1B | Contrastive Captioner on ViT-G/14 | 1 | 1 | |
| 533 | CodeBERT+AdvFusion | University of Leicester | 125M | transformer | 1 | 1 |
| 534 | CodeT5+ 2B | Salesforce | Unknown | T5-based encoder-decoder | 1 | 1 |
| 535 | CodeTrans-MT-Large | Unknown | Unknown | Unknown | 1 | 1 |
| 536 | Co-DETR (Swin-L) | Research | Unknown | Collaborative DETR + Swin-L backbone | 1 | 1 |
| 537 | Co-DETR (Swin-L) | — | — | — | 1 | 1 |
| 538 | Co-DETR (Swin-L) | Research | — | Transformer Detector | 1 | 1 |
| 539 | Co-DINO-Deformable-DETR++ (Swin-L, 36 epochs) | — | — | — | 1 | 1 |
| 540 | Co-DINO (ViT-L) | Sensetime / Sense-X | ~600M | DINO transformer detector with ViT-L backbone and collaborative hybrid assignment training | 1 | 1 |
| 541 | convnext_base.fb_in22k_ft_in1k | Meta | — | ConvNeXt-B, IN22K pre-train, IN1K fine-tune | 1 | 1 |
| 542 | ConvNeXt V2 Base | Meta | 89M | CNN | 1 | 1 |
| 543 | ConvNeXt V2 Tiny | Meta | 28M | CNN | 1 | 1 |
| 544 | ConvStem | Unknown | Unknown | Unknown | 1 | 1 |
| 545 | ConvTextTM | Unknown | Unknown | Unknown | 1 | 1 |
| 546 | CoTexT | Case Western Reserve University | — | Transformer encoder-decoder | 1 | 1 |
| 547 | Cross-Modal | Unknown | Unknown | Unknown | 1 | 1 |
| 548 | CUTeOCR | CUHK / HIT | Unknown | Scene text detector | 1 | 1 |
| 549 | CW_Detection | Independent | — | — | 1 | 1 |
| 550 | DAL | Unknown | Unknown | Unknown | 1 | 1 |