OpenAI: o4 Mini High wins on more metrics (4 of 6), but the right pick depends on what you optimize for — see the breakdown below.
| Metric | Anthropic: Claude Sonnet 4 | OpenAI: o4 Mini High |
|---|---|---|
| Intelligence Index | 25.5 | 25.6✓ |
| Coding Index | — | — |
| GPQA Diamond | 68% | 78%✓ |
| Design Arena Elo | 1197✓ | — |
| Speed (tokens/sec) | — | — |
| Latency | — | — |
| Input price /M | $3.00 | $1.10✓ |
| Output price /M | $15.00 | $4.40✓ |
| Context window | 1M✓ | 200K |
| Capabilities | ReasoningToolsVision | ReasoningToolsJSONVision |
OpenAI: o4 Mini High wins on more metrics (4 of 6), but the right pick depends on what you optimize for — see the breakdown below.
The table on this page compares Anthropic: Claude Sonnet 4 and OpenAI: o4 Mini High metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.