deepseek/deepseek-v4-flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| DeepInfrafp4 | $0.090 | $0.180 | 1.0M | 99.2% |
| StreamLakefp8 | $0.092 | $0.185 | 1.0M | 99.7% |
| GMICloudfp8 | $0.094 | $0.188 | 1.0M | 99.6% |
| AkashMLfp8 | $0.098 | $0.196 | 1.0M | 87.1% |
| DigitalOcean | $0.112 | $0.224 | 262K | 99.1% |
| SiliconFlowfp8 | $0.130 | $0.280 | 1.0M | 99.2% |
| Alibabafp8 | $0.134 | $0.268 | 1M | 99.9% |
| Venice | $0.138 | $0.275 | 1M | 99.5% |
| Morph | $0.139 | $0.278 | 1.0M | 99.6% |
| Io Netfp8 | $0.139 | $0.278 | 33K | 99.9% |
| Ionstreamfp4 | $0.140 | $0.280 | 1.0M | 91.5% |
| Parasailfp8 | $0.140 | $0.280 | 1.0M | 99.3% |
| Fireworks | $0.140 | $0.280 | 1.0M | 98.2% |
| Novitafp8 | $0.140 | $0.280 | 1.0M | 99.9% |
| Ambientfp4 | $0.140 | $0.280 | 1.0M | 83.6% |
| Cloudflare | $0.140 | $0.280 | 384K | 100% |
| AtlasCloudfp4 | $0.140 | $0.280 | 1.0M | 99.8% |
| DeepSeekcache | $0.140 | $0.280 | 1.0M | 100% |
| Baidufp8 | $0.140 | $0.280 | 1.0M | 99.6% |
| CoreWeavefp8 | $0.140 | $0.280 | 1.0M | 99.6% |
| Mancer 2fp4 | $0.200 | $0.500 | 1.0M | 98.3% |
DeepSeek V4 Flash costs $0.140 per million input tokens and $0.280 per million output tokens via OpenRouter, making it 72nd cheapest of 332 paid models.
DeepSeek V4 Flash scores 40.3 on the Artificial Analysis Intelligence Index, ranking 40th of 182 benchmarked models, with a GPQA Diamond score of 89%.
DeepSeek V4 Flash supports a 1.0M-token context window and can output up to 393K tokens. It accepts text input.