openai/gpt-oss-20b
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| CoreWeavefp4 | $0.030 | $0.130 | 131K | 99.9% |
| DeepInfrabf16 | $0.030 | $0.140 | 131K | 99.9% |
| Parasailfp4 | $0.030 | $0.150 | 131K | 78.8% |
| Novitafp4 | $0.040 | $0.150 | 131K | 100% |
| Phala | $0.040 | $0.150 | 131K | 99.9% |
| SiliconFlowfp8 | $0.040 | $0.180 | 131K | 93.9% |
| Together | $0.050 | $0.200 | 131K | — |
| Amazon Bedrock | $0.070 | $0.150 | 131K | — |
| Amazon Bedrock | $0.070 | $0.150 | 131K | 99.9% |
| $0.070 | $0.250 | 131K | 73.1% | |
| Fireworks | $0.070 | $0.300 | 131K | — |
| Groq | $0.075 | $0.300 | 131K | 99.3% |
gpt-oss-20b costs $0.030 per million input tokens and $0.130 per million output tokens via OpenRouter, making it 8th cheapest of 332 paid models.
gpt-oss-20b scores 14.9 on the Artificial Analysis Intelligence Index, ranking 136th of 182 benchmarked models, with a GPQA Diamond score of 69%.
gpt-oss-20b supports a 131K-token context window and can output up to 131K tokens. It accepts text input.