google/gemini-3.1-flash-lite
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| Googlecache | $0.250 | $1.50 | 1.0M | 98.2% |
| Googlecache | $0.125 | $0.750 | 1.0M | 98.2% |
| Googlecache | $0.450 | $2.70 | 1.0M | 98.2% |
| Google AI Studiocache | $0.250 | $1.50 | 1.0M | 98.8% |
| Google AI Studiocache | $0.125 | $0.750 | 1.0M | 98.8% |
| Google AI Studiocache | $0.450 | $2.70 | 1.0M | 98.8% |
Gemini 3.1 Flash Lite costs $0.250 per million input tokens and $1.50 per million output tokens via OpenRouter, making it 110th cheapest of 332 paid models.
Gemini 3.1 Flash Lite scores 25.0 on the Artificial Analysis Intelligence Index, ranking 105th of 182 benchmarked models, with a GPQA Diamond score of 82%.
Gemini 3.1 Flash Lite supports a 1.0M-token context window and can output up to 66K tokens. It accepts text, image, video, file, audio input.