Comparison
Llama 4 Scout vs Mistral Small 4
Mistral Small 4 charges 33% more per input token than Llama 4 Scout. Output pricing favors Llama 4 Scout, 50% less per token than Mistral Small 4.
| Field | Llama 4 Scout | Mistral Small 4 |
|---|---|---|
| Provider | Meta | Mistral AI |
| Context window | 10M | — |
| Max output | — | — |
| Input / 1M | $0.1033% cheaper | $0.15 |
| Output / 1M | $0.3050% cheaper | $0.60 |
| Cached input / 1M | — | — |
| Batch discount | — | 50% |
| Release date | 2025-04-05 | — |
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Llama 4 Scout | Mistral Small 4 |
|---|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $14.00cheaper | $24.00 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $24.30cheaper | $39.60 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $12.25cheaper | $19.50 |
Llama 4 Scout vs Mistral Small 4 FAQ
- Is Llama 4 Scout cheaper than Mistral Small 4?
- Llama 4 Scout costs $0.10 per 1M input tokens versus $0.15 for Mistral Small 4 — Llama 4 Scout is 33% cheaper per input token.
- Which is cheaper for output tokens, Llama 4 Scout or Mistral Small 4?
- Llama 4 Scout charges $0.30 per 1M output tokens; Mistral Small 4 charges $0.60 per 1M output tokens.
- Do Llama 4 Scout and Mistral Small 4 offer cache pricing?
- Llama 4 Scout has no documented cache discount. Mistral Small 4 has no documented cache discount.
Looking for another matchup? See all comparisons.