modelgrep

Most-used AI models for shell execution

Match · Updated 2026-08-04

DeepSeek V4 Pro is the most-used model for shell execution on OpenRouter, handling 24.1% of requests classified as shell execution over the last 7 days.

1.2%of all requests
2.9%of all tokens
7dwindow

This ranks models by how often they are actually chosen for shell execution in production traffic, not by benchmark score. Adoption reflects price, latency and availability as much as raw capability — which is exactly why it often disagrees with the benchmark leaderboards, and why it is worth reading alongside them.

#ModelRequest shareToken shareInput /MContext
1deepseek logoDeepSeek V4 Pro24.1%20.9%$0.4351.0M
2deepseek logoDeepSeek V4 Flash 04239.5%7.7%$0.1401.0M
3stepfun logoStep 3.7 Flash8.3%7.3%$0.200262K
4z-ai logoGLM 5.26.3%5.9%$0.7601.0M
5xiaomi logoMiMo-V2.55.4%8.5%$0.1401.1M
6tencent logoHy35.1%5.5%$0.132262K
7deepseek logoDeepSeek V4 Flash 04233.8%5.8%$0.1401.0M
8nemotron-3-ultra-550b-a55b-20260604:free2.9%4.4%
9openai logoGPT-5.6 Luna2.7%4.7%$0.1001.1M
10anthropic logoClaude Sonnet 4.62.4%1.2%$3.001M

Reading the two columns together: MiMo-V2.5, DeepSeek V4 Flash 0423, nemotron-3-ultra-550b-a55b-20260604:free take a noticeably larger share of tokens than of requests — meaning they are being used for the longer, heavier shell execution jobs rather than quick one-shot calls.

Frequently asked

What is the most-used AI model for shell execution?

DeepSeek V4 Pro is the most-used model for shell execution on OpenRouter, handling 24.1% of requests classified as shell execution over the last 7 days. It is followed by DeepSeek V4 Flash 0423 (9.5%) and Step 3.7 Flash (8.3%).

Is this a benchmark ranking?

No. This ranks by how often each model is actually chosen for shell execution in production traffic through OpenRouter — real-world adoption, which reflects price and availability as much as capability. For capability-based rankings, see the benchmark leaderboards.

How much of all AI traffic is shell execution?

Shell Execution accounts for 1.2% of classified requests and 2.9% of classified tokens on OpenRouter over the trailing 7 days.

Other tasks

Source: OpenRouter (openrouter.ai/rankings), as of 2026-08-04. Shares are of classified, sampled traffic over a trailing 7-day window; absolute volumes are not published.