Comparison
Claude Sonnet 5 vs Llama 4 Maverick
Per input token, Llama 4 Maverick runs 93% cheaper than Claude Sonnet 5. Llama 4 Maverick is 94% cheaper on output tokens. Claude Sonnet 5 offers a cheaper cached-input tier for repeated context.
| Field | Claude Sonnet 5 | Llama 4 Maverick |
|---|---|---|
| Provider | Anthropic | Meta |
| Context window | 1M | 1M |
| Max output | 128K | — |
| Input / 1M | $2.00 | $0.1593% cheaper |
| Output / 1M | $10.00 | $0.6094% cheaper |
| Cached input / 1M | $0.20 | — |
| Batch discount | 50% | — |
| Release date | — | — |
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Claude Sonnet 5 | Llama 4 Maverick |
|---|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $360.00 | $24.00cheaper |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $570.00 | $39.60cheaper |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $275.00 | $19.50cheaper |
Claude Sonnet 5 vs Llama 4 Maverick FAQ
- Is Claude Sonnet 5 cheaper than Llama 4 Maverick?
- Claude Sonnet 5 costs $2.00 per 1M input tokens versus $0.15 for Llama 4 Maverick — Llama 4 Maverick is 1233% cheaper per input token.
- Which is cheaper for output tokens, Claude Sonnet 5 or Llama 4 Maverick?
- Claude Sonnet 5 charges $10.00 per 1M output tokens; Llama 4 Maverick charges $0.60 per 1M output tokens.
- Which has the larger context window, Claude Sonnet 5 or Llama 4 Maverick?
- Claude Sonnet 5 does, with a 1,000,000-token context window.
- Do Claude Sonnet 5 and Llama 4 Maverick offer cache pricing?
- Claude Sonnet 5 discounts cached input to $0.20 per 1M tokens. Llama 4 Maverick has no documented cache discount.
Looking for another matchup? See all comparisons.