Comparison

Llama 4 Scout vs Mistral Small 4

Mistral Small 4 charges 33% more per input token than Llama 4 Scout. Output pricing favors Llama 4 Scout, 50% less per token than Mistral Small 4.

FieldLlama 4 Scout
Mistral Small 4
ProviderMetaMistral AI
Context window10M
Max output
Input / 1M$0.1033% cheaper$0.15
Output / 1M$0.3050% cheaper$0.60
Cached input / 1M
Batch discount50%
Release date2025-04-05

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadLlama 4 ScoutMistral Small 4

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$14.00cheaper$24.00

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$24.30cheaper$39.60

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$12.25cheaper$19.50

Llama 4 Scout vs Mistral Small 4 FAQ

Is Llama 4 Scout cheaper than Mistral Small 4?
Llama 4 Scout costs $0.10 per 1M input tokens versus $0.15 for Mistral Small 4 — Llama 4 Scout is 33% cheaper per input token.
Which is cheaper for output tokens, Llama 4 Scout or Mistral Small 4?
Llama 4 Scout charges $0.30 per 1M output tokens; Mistral Small 4 charges $0.60 per 1M output tokens.
Do Llama 4 Scout and Mistral Small 4 offer cache pricing?
Llama 4 Scout has no documented cache discount. Mistral Small 4 has no documented cache discount.

Looking for another matchup? See all comparisons.