meta-llama/llama-3.1-8b-instruct
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| DeepInfrafp8 | $0.020 | $0.040 | 131K | 99.9% |
| Novitafp8 | $0.020 | $0.050 | 16K | 99.8% |
| Groq | $0.050 | $0.080 | 131K | 99.8% |
| Cloudflarefp8 | $0.152 | $0.287 | 32K | 99.8% |
| CoreWeavebf16 | $0.220 | $0.220 | 128K | 100% |
Llama 3.1 8B Instruct costs $0.050 per million input tokens and $0.080 per million output tokens via OpenRouter, making it 22nd cheapest of 332 paid models.
Llama 3.1 8B Instruct supports a 131K-token context window and can output up to 131K tokens. It accepts text input.