Meta

Llama 4 Maverick API Pricing

Aggregator pricing; Meta does not sell a first-party priced API (llama.developer.meta.com is waitlist-only with no published pricing). Mixture-of-experts model, 17B active / ~400B total parameters, 128 experts.

Verified against Meta’s pricing page on 2026-07-02.

View source ↗

Facts

Specs

  • FamilyLlama 4
  • API IDmeta-llama/llama-4-maverick
  • ModalityMultimodal
  • Context window1M tokens
  • Max output
  • Release dateNot published

Pricing (per 1M tokens)

  • Input$0.15 / 1M tokens
  • Output$0.60 / 1M tokens
  • Cached input
  • Cache write
  • Batch discount

Capabilities

vision, tool-use, mixture-of-experts

Notes

Aggregator pricing; Meta does not sell a first-party priced API (llama.developer.meta.com is waitlist-only with no published pricing). Mixture-of-experts model, 17B active / ~400B total parameters, 128 experts.

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadEst. monthly cost

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$24.00

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$39.60

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$19.50

Price history

Changes we have documented since we began tracking

No price changes recorded since we began tracking.

Llama 4 Maverick pricing FAQ

How much does Llama 4 Maverick cost per 1M tokens?
Llama 4 Maverick costs $0.15 per 1M input tokens and $0.60 per 1M output tokens.
How much does Llama 4 Maverick cost per request?
A typical request of 800 input and 200 output tokens costs about $0.00024 on Llama 4 Maverick.
Is Llama 4 Maverick cheaper than Llama 4 Scout?
Llama 4 Maverick is more expensive than Llama 4 Scout on input tokens, by about 50%.
Does Llama 4 Maverick have cache pricing?
No cache discount is documented for Llama 4 Maverick on the provider's pricing page.