Rankings
Best Value per Context Token
Models offering the most context window per dollar of input cost.
Models ranked
1283
Top context
10M
Llama 4 Scout 17b 16e
Leading provider
Openai
241 models in this ranking
Leaderboard
Tap any row to see full specs and comparisons.
| Rank | Model | Context | Max output | Input $/M | Value score |
|---|---|---|---|---|---|
| 1 | Llama 4 Scout 17b 16eMeta | 10M | 16K | $0.050/M | 10000K / $0.05 |
| 2 | Llama 4 Scout 17b 128e Instruct MaasMeta | 10M | 10M | $0.250/M | 10000K / $0.25 |
| 3 | Llama 4 Scout 17b 16e Instruct MaasMeta | 10M | 10M | $0.250/M | 10000K / $0.25 |
| 4 | Openai Gpt 5 NanoOpenai | 5M | 16K | $0.150/M | 5000K / $0.15 |
| 5 | Qwen3.7 FlashAlibaba | 1M | 66K | $0.030/M | 1000K / $0.03 |
| 6 | DeepSeek V4 Flash LatestDeepseek | 1.3M | 393K | $0.040/M | 1310K / $0.04 |
| 7 | DeepSeek V4 Flash 0731Deepseek | 1.0M | 384K | $0.040/M | 1048K / $0.04 |
| 8 | DeepSeek V4 FlashDeepseek | 1.0M | 384K | $0.300/M | 1048K / $0.05 |
| 9 | GPT-6 Luna Pro (batch)Openai | 1.1M | 128K | $0.050/M | 1050K / $0.05 |
| 10 | GPT-6 Luna (batch)Openai | 1.1M | 128K | $0.050/M | 1050K / $0.05 |
| 11 | Gemini 2.5 Flash Lite (batch)Google | 1.0M | 66K | $0.050/M | 1048K / $0.05 |
| 12 | GPT-4.1 Nano (batch)Openai | 1.0M | 33K | $0.050/M | 1047K / $0.05 |
| 13 | Llama 4 Maverick 17b 128e Instruct Fp8Meta | 1M | 16K | $0.050/M | 1000K / $0.05 |
| 14 | Qwen Turbo LatestAlibaba | 1M | 16K | $0.050/M | 1000K / $0.05 |
| 15 | Qwen Turbo 2024 11 01Alibaba | 1M | 8K | $0.050/M | 1000K / $0.05 |
| 16 | Qwen Turbo 2025 04 28Alibaba | 1M | 16K | $0.050/M | 1000K / $0.05 |
| 17 | GLM 5.3 Flash (batch)Z Ai | 1.0M | 944K | $0.060/M | 1048K / $0.06 |
| 18 | GPT-5 Nano (batch)Openai | 400K | 128K | $0.025/M | 400K / $0.03 |
| 19 | Qwen3.5-FlashAlibaba | 1M | 66K | $0.065/M | 1000K / $0.07 |
| 20 | Gemini 2 0 Flash LiteGoogle | 1.0M | 8K | $0.075/M | 1048K / $0.07 |
| 21 | Gemini 2.0 Flash LiteGoogle | 1.0M | 8K | $0.075/M | 1048K / $0.07 |
| 22 | Qwen3.8 FlashAlibaba | 1M | 131K | $0.090/M | 1000K / $0.09 |
| 23 | GPT-5.6 Luna Pro (batch)Openai | 1.1M | 128K | $0.100/M | 1050K / $0.10 |
| 24 | GPT-5.6 Luna (batch)Openai | 1.1M | 128K | $0.100/M | 1050K / $0.10 |
| 25 | GPT Luna LatestOpenai | 1.1M | 128K | $0.100/M | 1050K / $0.10 |
Showing 25 of 1283 models
More rankings
Explore other model leaderboards.
Largest Context Window
Models ranked by maximum context window size.
Best for RAG
Models best suited for Retrieval-Augmented Generation workloads.
Best for AI Agents
Models with the capabilities needed to power autonomous agent workflows.
Best for Document Processing
Models with large enough context to process long documents in one pass.
Best Multimodal
Vision-capable models with the largest context windows.
Best Reasoning
Models with extended thinking or strong reasoning capabilities.
Best for Chatbots
Fast, cost-efficient models ideal for real-time conversational applications.