DeepSeek V4 Pro is the most-used model for shell execution on OpenRouter, handling 24.1% of requests classified as shell execution over the last 7 days.
This ranks models by how often they are actually chosen for shell execution in production traffic, not by benchmark score. Adoption reflects price, latency and availability as much as raw capability — which is exactly why it often disagrees with the benchmark leaderboards, and why it is worth reading alongside them.
| # | Model | Request share | Token share | Input /M | Context |
|---|---|---|---|---|---|
| 1 | 24.1% | 20.9% | $0.435 | 1.0M | |
| 2 | 9.5% | 7.7% | $0.140 | 1.0M | |
| 3 | 8.3% | 7.3% | $0.200 | 262K | |
| 4 | 6.3% | 5.9% | $0.760 | 1.0M | |
| 5 | 5.4% | 8.5% | $0.140 | 1.1M | |
| 6 | 5.1% | 5.5% | $0.132 | 262K | |
| 7 | 3.8% | 5.8% | $0.140 | 1.0M | |
| 8 | nemotron-3-ultra-550b-a55b-20260604:free | 2.9% | 4.4% | — | — |
| 9 | 2.7% | 4.7% | $0.100 | 1.1M | |
| 10 | 2.4% | 1.2% | $3.00 | 1M |
Reading the two columns together: MiMo-V2.5, DeepSeek V4 Flash 0423, nemotron-3-ultra-550b-a55b-20260604:free take a noticeably larger share of tokens than of requests — meaning they are being used for the longer, heavier shell execution jobs rather than quick one-shot calls.
DeepSeek V4 Pro is the most-used model for shell execution on OpenRouter, handling 24.1% of requests classified as shell execution over the last 7 days. It is followed by DeepSeek V4 Flash 0423 (9.5%) and Step 3.7 Flash (8.3%).
No. This ranks by how often each model is actually chosen for shell execution in production traffic through OpenRouter — real-world adoption, which reflects price and availability as much as capability. For capability-based rankings, see the benchmark leaderboards.
Shell Execution accounts for 1.2% of classified requests and 2.9% of classified tokens on OpenRouter over the trailing 7 days.
Source: OpenRouter (openrouter.ai/rankings), as of 2026-08-04. Shares are of classified, sampled traffic over a trailing 7-day window; absolute volumes are not published.