modelgrep

Anthropic: Claude Opus 4 vs Google: Gemini 3.1 Flash Lite

Google: Gemini 3.1 Flash Lite wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

MetricAnthropic: Claude Opus 4Google: Gemini 3.1 Flash Lite
Intelligence Index25.525.0
Coding Index34.7
GPQA Diamond70%82%
Design Arena Elo1196
Speed (tokens/sec)
Latency
Input price /M$15.00$0.250
Output price /M$75.00$1.50
Context window200K1.0M
CapabilitiesReasoningToolsVisionReasoningToolsJSONVisionAudio

Frequently asked

Is Anthropic: Claude Opus 4 better than Google: Gemini 3.1 Flash Lite?

Google: Gemini 3.1 Flash Lite wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between Anthropic: Claude Opus 4 and Google: Gemini 3.1 Flash Lite?

The table on this page compares Anthropic: Claude Opus 4 and Google: Gemini 3.1 Flash Lite metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.