modelgrep

DeepSeek: DeepSeek V3.1 Terminus vs Inception: Mercury 2

Inception: Mercury 2 wins on more metrics (5 of 6), but the right pick depends on what you optimize for — see the breakdown below.

MetricDeepSeek: DeepSeek V3.1 TerminusInception: Mercury 2
Intelligence Index21.421.4
Coding Index31.1
GPQA Diamond75%77%
Design Arena Elo1042
Speed (tokens/sec)
Latency
Input price /M$0.270$0.250
Output price /M$1.00$0.750
Context window164K128K
CapabilitiesReasoningToolsJSONReasoningToolsJSON

Frequently asked

Is DeepSeek: DeepSeek V3.1 Terminus better than Inception: Mercury 2?

Inception: Mercury 2 wins on more metrics (5 of 6), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between DeepSeek: DeepSeek V3.1 Terminus and Inception: Mercury 2?

The table on this page compares DeepSeek: DeepSeek V3.1 Terminus and Inception: Mercury 2 metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.