modelgrep

Anthropic: Claude Opus 4.8 vs Google: Gemini 3.1 Pro Preview

Google: Gemini 3.1 Pro Preview wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

MetricAnthropic: Claude Opus 4.8Google: Gemini 3.1 Pro Preview
Intelligence Index55.746.5
Coding Index74.368.8
GPQA Diamond92%94%
Design Arena Elo12781335
Speed (tokens/sec)
Latency
Input price /M$5.00$2.00
Output price /M$25.00$12.00
Context window1M1.0M
CapabilitiesReasoningToolsJSONVisionReasoningToolsJSONVisionAudio

Frequently asked

Is Anthropic: Claude Opus 4.8 better than Google: Gemini 3.1 Pro Preview?

Google: Gemini 3.1 Pro Preview wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between Anthropic: Claude Opus 4.8 and Google: Gemini 3.1 Pro Preview?

The table on this page compares Anthropic: Claude Opus 4.8 and Google: Gemini 3.1 Pro Preview metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.