google/gemini-3.5-flash-lite
Gemini 3.5 Flash-Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| Google AI Studiocache | $0.300 | $2.50 | 1.0M | 99.4% |
| Google AI Studiocache | $0.150 | $1.25 | 1.0M | 99.4% |
| Google AI Studiocache | $0.540 | $4.50 | 1.0M | 99.4% |
| Googlecache | $0.300 | $2.50 | 1.0M | 100% |
| Googlecache | $0.150 | $1.25 | 1.0M | 100% |
| Googlecache | $0.540 | $4.50 | 1.0M | 100% |
Gemini 3.5 Flash-Lite costs $0.300 per million input tokens and $2.50 per million output tokens via OpenRouter, making it 123rd cheapest of 310 paid models.
Gemini 3.5 Flash-Lite scores 36.5 on the Artificial Analysis Intelligence Index, ranking 45th of 155 benchmarked models, with a GPQA Diamond score of 84%.
Gemini 3.5 Flash-Lite generates around 82 tokens per second with 873ms time-to-first-token (p50), the 12th fastest tracked model.
Gemini 3.5 Flash-Lite supports a 1.0M-token context window and can output up to 66K tokens. It accepts text, image, video, file, audio input.