Rankings
Best for RAG
Models best suited for Retrieval-Augmented Generation workloads.
Models ranked
1152
Top context
131K
Llama3 2 11b Vision
Leading provider
Openai
234 models in this ranking
Leaderboard
Tap any row to see full specs and comparisons.
| Rank | Model | Context | Max output | Input $/M | RAG score |
|---|---|---|---|---|---|
| 1 | Llama3 2 11b VisionMeta | 131K | 131K | $0.015/M | 131K ctx |
| 2 | Llama3 2 3bMeta | 131K | 131K | $0.015/M | 131K ctx |
| 3 | Granite 4.0 MicroIbm | 131K | 131K | $0.017/M | 131K ctx |
| 4 | gpt-oss-20bOpenai | 131K | 131K | $0.018/M | 131K ctx |
| 5 | Mistral NemoMistral | 131K | 4K | $0.019/M | 131K ctx |
| 6 | Mistral Nemo Instruct 2407Mistral | 131K | 131K | $0.019/M | 131K ctx |
| 7 | Llama 3 2 3bMeta | 131K | 131K | $0.020/M | 131K ctx |
| 8 | Meta Llama 3 1 8b InstructMeta | 131K | 131K | $0.020/M | 131K ctx |
| 9 | Gemma 4 E4b ItGoogle | 131K | — | $0.020/M | 131K ctx |
| 10 | Llama 3 2 1bMeta | 128K | 128K | $0.020/M | 128K ctx |
| 11 | Llama 3 1 8bMeta | 131K | 16K | $0.020/M | 131K ctx |
| 12 | Llama Guard 3 8BMeta | 131K | 131K | $0.020/M | 131K ctx |
| 13 | Qwen2 Vl 7bAlibaba | 131K | 131K | $0.020/M | 131K ctx |
| 14 | Meta Llama 3 1 8bMeta | 128K | 2K | $0.020/M | 128K ctx |
| 15 | GPT-5 Nano (batch)Openai | 400K | 128K | $0.025/M | 400K ctx |
| 16 | Qwen3.7 FlashAlibaba | 1M | 66K | $0.030/M | 1000K ctx |
| 17 | DeepSeek V4 Flash 0731Deepseek | 1.0M | 384K | $0.030/M | 1048K ctx |
| 18 | Hermes3 8bNous Research | 131K | 131K | $0.025/M | 131K ctx |
| 19 | Llama3 1 8bMeta | 128K | 128K | $0.025/M | 128K ctx |
| 20 | gpt-oss-20b (batch)Openai | 131K | 118K | $0.024/M | 131K ctx |
| 21 | Deepseek R1 Distill Llama 8bMeta | 131K | — | $0.025/M | 131K ctx |
| 22 | DeepSeek V4 Flash LatestDeepseek | 1.3M | 393K | $0.040/M | 1310K ctx |
| 23 | gpt-oss-120b (batch)Openai | 131K | 118K | $0.030/M | 131K ctx |
| 24 | GLM 5.3 FlashZ Ai | 1.0M | 131K | $0.150/M | 1048K ctx |
| 25 | Qwen3 4b Fp8Alibaba | 128K | 20K | $0.030/M | 128K ctx |
Showing 25 of 1152 models
More rankings
Explore other model leaderboards.
Largest Context Window
Models ranked by maximum context window size.
Best for AI Agents
Models with the capabilities needed to power autonomous agent workflows.
Best for Document Processing
Models with large enough context to process long documents in one pass.
Best Value per Context Token
Models offering the most context window per dollar of input cost.
Best Multimodal
Vision-capable models with the largest context windows.
Best Reasoning
Models with extended thinking or strong reasoning capabilities.
Best for Chatbots
Fast, cost-efficient models ideal for real-time conversational applications.