modelgrep

Anthropic: Claude Haiku 4.5 vs OpenAI: gpt-oss-120b

OpenAI: gpt-oss-120b wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

MetricAnthropic: Claude Haiku 4.5OpenAI: gpt-oss-120b
Intelligence Index23.723.8
Coding Index30.4
GPQA Diamond65%78%
Design Arena Elo11531029
Speed (tokens/sec)
Latency
Input price /M$1.00$0.037
Output price /M$5.00$0.170
Context window200K131K
CapabilitiesReasoningToolsJSONVisionReasoningToolsJSON

Frequently asked

Is Anthropic: Claude Haiku 4.5 better than OpenAI: gpt-oss-120b?

OpenAI: gpt-oss-120b wins on more metrics (5 of 7), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between Anthropic: Claude Haiku 4.5 and OpenAI: gpt-oss-120b?

The table on this page compares Anthropic: Claude Haiku 4.5 and OpenAI: gpt-oss-120b metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.