modelgrep
meta-llama logo

Meta: Llama 4 Maverick

meta-llama/llama-4-maverick

141st smartest of 182ToolsJSONVision
Use via OpenRouter ↗
Intelligence
14.3
141st of 182
Design Elo
957
3D
Speed
tokens/sec
Latency
first token
Input price
$0.200
102nd cheapest
Context
1.0M
16K max out

How it compares

Smarter than23%
of all ranked models
Cheaper than69%
of all ranked models

Overview

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Benchmarks

independent · Artificial Analysis & Design Arena
Artificial Analysis48th percentile
Intelligence Index
14.3
Coding Index
16.3
GPQA Diamond
67%
Humanity's Last Exam
5%
SciCode
33%
Tau²-Bench (agentic)
18%
Design Arena · Elo164 tournaments
3D
957
Data Viz
912
Website
895

Providers & pricing (5)

ProviderIn $/MOut $/MUptime
DeepInfrafp8$0.200$0.80099.8%
DigitalOcean$0.250$0.87099.9%
Novitafp8$0.270$0.85099.5%
Parasailfp8$0.350$1.0099.7%
Google$0.350$1.15100%

Specifications

Context window1.0M
Max output16K
Knowledge cutoffAug 2024
Input modalitiestext, image
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Llama 4 Maverick FAQ

How much does Llama 4 Maverick cost?

Llama 4 Maverick costs $0.200 per million input tokens and $0.800 per million output tokens via OpenRouter, making it 102nd cheapest of 332 paid models.

How smart is Llama 4 Maverick?

Llama 4 Maverick scores 14.3 on the Artificial Analysis Intelligence Index, ranking 141st of 182 benchmarked models, with a GPQA Diamond score of 67%.

What is Llama 4 Maverick's context window?

Llama 4 Maverick supports a 1.0M-token context window and can output up to 16K tokens. It accepts text, image input.

Compare head-to-head