The cheapest Google model is Gemini 2.5 Flash Lite (batch) at $0.050 per million input tokens. Gemma 3 4B ($0.050) and Gemma 3 12B ($0.050) round out the top three.
AI models ranked by input token price. The cheapest LLM APIs and most affordable AI models, from budget open-weight models to discounted frontier models.
The cheapest Google model is Gemini 2.5 Flash Lite (batch) at $0.050 per million input tokens. Gemma 3 4B ($0.050) and Gemma 3 12B ($0.050) round out the top three.
Gemma 3 4B ($0.050) is the closest alternative on this metric, followed by Gemma 3 12B ($0.050). See the full ranking above for the tradeoffs.
modelgrep tracks 39 Google models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Gemini 3.5 Flash. 25 of them qualify for this ranking.