Qwen: Qwen3.5-9B wins on more metrics (5 of 5), but the right pick depends on what you optimize for — see the breakdown below.
| Metric | DeepSeek: DeepSeek V3.1 Terminus | Qwen: Qwen3.5-9B |
|---|---|---|
| Intelligence Index | 21.4 | 21.4 |
| Coding Index | — | 28.7✓ |
| GPQA Diamond | 75% | 81%✓ |
| Design Arena Elo | — | — |
| Speed (tokens/sec) | — | — |
| Latency | — | — |
| Input price /M | $0.270 | $0.100✓ |
| Output price /M | $1.00 | $0.150✓ |
| Context window | 164K | 262K✓ |
| Capabilities | ReasoningToolsJSON | ReasoningToolsJSONVision |
Qwen: Qwen3.5-9B wins on more metrics (5 of 5), but the right pick depends on what you optimize for — see the breakdown below.
The table on this page compares DeepSeek: DeepSeek V3.1 Terminus and Qwen: Qwen3.5-9B metric by metric — intelligence benchmarks, coding score, speed, latency, context window and per-token price — with the winner marked on each row. Data refreshes hourly.