Codesota · Models2,268 models indexed · 854 match filter
Editorial · Models

Models with recorded evidence.

Start with a research area, drill into a vendor, or page through the full index. Vendor aliases and legacy area IDs are grouped here; original model IDs and model links stay unchanged. Only models with at least one benchmark score appear — a model without a recorded score can’t be ranked.

Vendor:Areas overviewSpeakLeash · 263Alibaba · 104Google · 102OpenAI · 86Meta · 68Microsoft · 49Anthropic · 44DeepSeek · 34Mistral · 30mistralai · 19CYFRAGOVPL · 14NVIDIA · 14Zhipu AI · 13internlm · 10xAI · 10ByteDance · 9Baidu · 8ibm-granite · 8PLLuM · 8allenai · 7Amazon · 7MiniMax · 7Mistral AI · 7Remek · 7Shanghai AI Lab · 7utter-project · 7CohereForAI · 6Salesforce · 601-ai · 5Cohere · 5Moonshot AI · 5NousResearch · 5THUML · 5gguf-iq · 4IBM · 4Meituan · 4openchat · 4Stanford · 4THUDM · 4tiiuae · 4UC San Diego · 4VikParuchuri · 4Allen AI · 3BAAI · 3Du et al. · 3ForgeCode · 3Fudan University · 3gguf · 3gguf11bv30 · 3gguf7bv30 · 3IDEA Research · 3Liao et al. · 3Moonshot.AI · 3Nam Tuan Ly / NII · 3OpenDataLab · 3OPI-PG · 3upstage · 3ViCoS Lab Ljubljana · 3Xiaomi · 3Zhao et al. · 3+ 243 smaller vendors (288 models)
§ 01 · Natural Language Processing models

854 models in Natural Language Processing · page 16 of 18.

#ModelVendorParametersArchitectureBenchmarksResults
751OLMo-2-7B-1124 (olmOCR-peS2o)———55
752openai/gpt-oss-120b (API)OpenAI120B—15
753Qwen/Qwen2.5-0.5B-InstructAlibaba0.49B—15
754Qwen/Qwen3-14B non-thinking (API)Alibaba14B—15
755Qwen/Qwen3-235B-A22B non-thinking (API)Alibaba235B—15
756Qwen/Qwen3-30B-A3B non-thinking (API)Alibaba30B—15
757Qwen/Qwen3-32B non-thinking (API)Alibaba32B—15
758Qwen/Qwen3.5-27B non-thinking (API)Alibaba27B—15
759Qwen/Qwen3.5-27B thinking (API)Alibaba27B—15
760Qwen/Qwen3.5-35B-A3B non-thinking (API)Alibaba35B—15
761Qwen/Qwen3.5-35B-A3B thinking (API)Alibaba35B—15
762Qwen/Qwen3.5-9B non-thinking (API, FP8)Alibaba9B—15
763Qwen/Qwen3-8B non-thinking (API)Alibaba8B—15
764speakleash/Bielik-Minitron-7B-v3.0-InstructSpeakLeash7.35B—15
765E5-Mistral-7B-instructMicrosoft7BMistral-7B (LLM-based embedding)34
766GTE-Qwen2-7B-instructAlibaba7BQwen2-7B (LLM-based embedding)34
767Helium———44
768T5-11BGoogleUnknownUnknown24
769Trinity Large PreviewArcee AI——44
770Claude 3.5 SonnetAnthropic—Transformer (LLM)33
771Gemini UltraGoogleUnknownTransformer (decoder-only)33
772GLM-5.1———33
773NV-Embed-v2NVIDIA7BMistral-7B (LLM-based embedding)23
774PEGASUS-LargeGoogleUnknownTransformer encoder-decoder (gap-sentence generation pre-training)13
775ALBERT ensemble———22
776BART———22
777ByT5 XXL———22
778ColBERTv2Stanford110MBERT (late interaction)22
779HunyuanOCR (1B)Unknown——22
780Mistral 7BMistral AIUnknownTransformer (decoder-only, GQA + sliding window attention)22
781ST-MoE-32BGoogle BrainUnknownSparse Mixture-of-Experts Transformer22
782ALBERT-xxlarge-v2Google235MALBERT-xxlarge11
783all-MiniLM-L6-v2Sentence-Transformers22MMiniLM-L6 (BERT-like)11
784BERT + AoAHIT & iFLYTEK——11
785BERT + ConvLSTM + MTL + Verifier (ensemble)Layer 6 AI——11
786BERT + DAE + AoA (single model)HIT & iFLYTEK——11
787BERT (Google AI)Google AI——11
788BERT Large———11
789DeBERTa (ensemble)Microsoft——11
790DeepLDeepL SE—Transformer (NMT)11
791embeddinggemma-300m———11
792Enhanced Albert+Verifier3 (ensemble)Microsoft STCA AIC——11
793ERNIE 3.0Baidu——11
794F2LLM-0.6B———11
795F2LLM-1.7B———11
796F2LLM-4B———11
797F2LLM-v2-0.6B———11
798F2LLM-v2-14B———11
799F2LLM-v2-1.7B———11
800F2LLM-v2-330M———11