Polarison

Meta · open weights

Llama 4 Maverick API pricing by provider

Llama 4 Maverick is available from 5 providers. The cheapest is DigitalOcean at $0.1875 input / $0.6525 output per 1M tokens. The priciest host, Google, charges 1.8x the cheapest.

Prices from the OpenRouter API, retrieved September 15, 2026. Weights: meta-llama/Llama-4-Maverick-17B-128E-Instruct.

ProviderPrecisionInput / 1MOutput / 1MDocument Q&A / monthContextMax output
DigitalOcean$0.1875$0.6525$94.58128K16K
DeepInfrafp8$0.20$0.80$104.001.05M16K
Novitafp8$0.27$0.85$133.501.05M8K
Parasailfp8$0.35$1.00$170.00524K33K
Google$0.35$1.15$174.50524K8K

Sorted by blended price (3 input : 1 output). Document Q&A assumes 50,000 questions a month with 8,000 input and 600 output tokens each. Precision is reported by the provider; “—” means not disclosed.

Frequently asked questions

What is the cheapest Llama 4 Maverick API provider?
DigitalOcean is the cheapest healthy provider we track at $0.1875 per 1M input tokens and $0.6525 per 1M output tokens.
How much does Llama 4 Maverick cost?
Across 5 providers, input prices range from $0.1875 to $0.35 per 1M tokens and output prices from $0.6525 to $1.15.
What is Llama 4 Maverick’s context window?
Llama 4 Maverick supports up to 1.05M tokens, but some providers serve a smaller context window — check the table before choosing.

Other open-weight models