modelgrep

Anthropic: Claude 3 Haiku vs Microsoft: Phi 4

Microsoft: Phi 4 wins on more metrics (4 of 5), but the right pick depends on what you optimize for — see the breakdown below.

MetricAnthropic: Claude 3 HaikuMicrosoft: Phi 4
Intelligence Index3.94.9
Coding Index
GPQA Diamond37%57%
Design Arena Elo
Speed (tokens/sec)
Latency
Input price /M$0.250$0.070
Output price /M$1.25$0.140
Context window200K16K
CapabilitiesToolsVisionJSON

Frequently asked

Is Anthropic: Claude 3 Haiku better than Microsoft: Phi 4?

Microsoft: Phi 4 wins on more metrics (4 of 5), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between Anthropic: Claude 3 Haiku and Microsoft: Phi 4?

The table on this page compares Anthropic: Claude 3 Haiku and Microsoft: Phi 4 metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.