modelgrep

Microsoft: Phi 4 vs Nous: Hermes 3 70B Instruct

Microsoft: Phi 4 wins on more metrics (3 of 5), but the right pick depends on what you optimize for — see the breakdown below.

MetricMicrosoft: Phi 4Nous: Hermes 3 70B Instruct
Intelligence Index4.95.1
Coding Index
GPQA Diamond57%40%
Design Arena Elo
Speed (tokens/sec)
Latency
Input price /M$0.070$0.700
Output price /M$0.140$0.700
Context window16K131K
CapabilitiesJSONJSON

Frequently asked

Is Microsoft: Phi 4 better than Nous: Hermes 3 70B Instruct?

Microsoft: Phi 4 wins on more metrics (3 of 5), but the right pick depends on what you optimize for — see the breakdown below.

What are the biggest differences between Microsoft: Phi 4 and Nous: Hermes 3 70B Instruct?

The table on this page compares Microsoft: Phi 4 and Nous: Hermes 3 70B Instruct metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.