meta-llama/llama-4-maverick
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| DeepInfrafp8 | $0.200 | $0.800 | 1.0M | 99.8% |
| DigitalOcean | $0.250 | $0.870 | 128K | 99.9% |
| Novitafp8 | $0.270 | $0.850 | 1.0M | 99.5% |
| Parasailfp8 | $0.350 | $1.00 | 524K | 99.7% |
| $0.350 | $1.15 | 524K | 100% |
Llama 4 Maverick costs $0.200 per million input tokens and $0.800 per million output tokens via OpenRouter, making it 102nd cheapest of 332 paid models.
Llama 4 Maverick scores 14.3 on the Artificial Analysis Intelligence Index, ranking 141st of 182 benchmarked models, with a GPQA Diamond score of 67%.
Llama 4 Maverick supports a 1.0M-token context window and can output up to 16K tokens. It accepts text, image input.