Nous: Hermes 3 70B Instruct wins on more metrics (3 of 5), but the right pick depends on what you optimize for — see the breakdown below.
| Metric | Anthropic: Claude 3 Haiku | Nous: Hermes 3 70B Instruct |
|---|---|---|
| Intelligence Index | 3.9 | 5.1✓ |
| Coding Index | — | — |
| GPQA Diamond | 37% | 40%✓ |
| Design Arena Elo | — | — |
| Speed (tokens/sec) | — | — |
| Latency | — | — |
| Input price /M | $0.250✓ | $0.700 |
| Output price /M | $1.25 | $0.700✓ |
| Context window | 200K✓ | 131K |
| Capabilities | ToolsVision | JSON |
Nous: Hermes 3 70B Instruct wins on more metrics (3 of 5), but the right pick depends on what you optimize for — see the breakdown below.
The table on this page compares Anthropic: Claude 3 Haiku and Nous: Hermes 3 70B Instruct metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.