DeepSeek V4 Flash 0423 is the most-used model for tool dispatch on OpenRouter, handling 22.5% of requests classified as tool dispatch over the last 7 days.
This ranks models by how often they are actually chosen for tool dispatch in production traffic, not by benchmark score. Adoption reflects price, latency and availability as much as raw capability — which is exactly why it often disagrees with the benchmark leaderboards, and why it is worth reading alongside them.
| # | Model | Request share | Token share | Input /M | Context |
|---|---|---|---|---|---|
| 1 | 22.5% | 19.2% | $0.140 | 1.0M | |
| 2 | 3.9% | 7.5% | $0.140 | 1.0M | |
| 3 | 3.9% | 6.2% | $0.760 | 1.0M | |
| 4 | 3.9% | 0.60% | $0.037 | 131K | |
| 5 | 3.7% | 5.2% | $3.00 | 1.0M | |
| 6 | 3.7% | 5.9% | $0.100 | 1.1M | |
| 7 | 3.4% | 6.7% | $0.132 | 262K | |
| 8 | 3.0% | 0.70% | $0.030 | 131K | |
| 9 | 2.8% | 4.4% | $0.435 | 1.0M | |
| 10 | 2.5% | 1.2% | $0.500 | 1.0M |
Reading the two columns together: DeepSeek V4 Flash 0423, GLM 5.2, GPT-5.6 Luna take a noticeably larger share of tokens than of requests — meaning they are being used for the longer, heavier tool dispatch jobs rather than quick one-shot calls.
DeepSeek V4 Flash 0423 is the most-used model for tool dispatch on OpenRouter, handling 22.5% of requests classified as tool dispatch over the last 7 days. It is followed by DeepSeek V4 Flash 0423 (3.9%) and GLM 5.2 (3.9%).
No. This ranks by how often each model is actually chosen for tool dispatch in production traffic through OpenRouter — real-world adoption, which reflects price and availability as much as capability. For capability-based rankings, see the benchmark leaderboards.
Tool Dispatch accounts for 0.70% of classified requests and 1.2% of classified tokens on OpenRouter over the trailing 7 days.
Source: OpenRouter (openrouter.ai/rankings), as of 2026-08-04. Shares are of classified, sampled traffic over a trailing 7-day window; absolute volumes are not published.