modelgrep

Longest-Context Google Models

Match · Updated July 2026

Gemini 3.6 Flash has the largest context window of any Google model, at 1.0M tokens. Gemini 3.6 Flash (batch) (1.0M) and Gemini 3.5 Flash Lite (1.0M) round out the top three.

1.0MContext
50.1Intelligence
$1.50Input /M

AI models with the largest context windows, ranked by token capacity. The best large language models for long documents, codebases and extended conversations.

  1. 1google logo
    gemini-3.6-flash
    ReasoningToolsJSON+250.1 intel · $1.50/M
    1.0M
    Context
  2. 2google logo
    gemini-3.6-flash:batch
    ReasoningToolsJSON+250.1 intel · $0.750/M
    1.0M
    Context
  3. 3google logo
    gemini-3.5-flash-lite
    ReasoningToolsJSON+236.5 intel · $0.300/M
    1.0M
    Context
  4. 4google logo
    gemini-3.5-flash-lite:batch
    ReasoningToolsJSON+236.5 intel · $0.150/M
    1.0M
    Context
  5. 5google logo
    gemini-3.5-flash
    ReasoningToolsJSON+250.2 intel · $1.50/M
    1.0M
    Context
  6. 6google logo
    gemini-3.5-flash:batch
    ReasoningToolsJSON+250.2 intel · $0.750/M
    1.0M
    Context
  7. 7google logo
    gemini-3.1-flash-lite
    ReasoningToolsJSON+225.0 intel · $0.250/M
    1.0M
    Context
  8. 8google logo
    gemini-3.1-flash-lite:batch
    ReasoningToolsJSON+225.0 intel · $0.125/M
    1.0M
    Context
  9. 9google logo
    lyria-3-pro-preview
    JSONVisionFree/M
    1.0M
    Context
  10. 10google logo
    lyria-3-clip-preview
    JSONVisionFree/M
    1.0M
    Context
  11. 11google logo
    gemini-3.1-flash-lite-preview
    ReasoningToolsJSON+225.0 intel · $0.250/M
    1.0M
    Context
  12. 12google logo
    gemini-3.1-pro-preview-customtools
    ReasoningToolsJSON+2$2.00/M
    1.0M
    Context
  13. 13google logo
    gemini-3.1-pro-preview
    ReasoningToolsJSON+246.5 intel · $2.00/M
    1.0M
    Context
  14. 14google logo
    gemini-3.1-pro-preview:batch
    ReasoningToolsJSON+246.5 intel · $1.00/M
    1.0M
    Context
  15. 15google logo
    gemini-3-flash-preview
    ReasoningToolsJSON+2$0.500/M
    1.0M
    Context
  16. 16google logo
    gemini-3-flash-preview:batch
    ReasoningToolsJSON+2$0.250/M
    1.0M
    Context
  17. 17google logo
    gemini-2.5-flash-lite
    ReasoningToolsJSON+26.9 intel · $0.100/M
    1.0M
    Context
  18. 18google logo
    gemini-2.5-flash-lite:batch
    ReasoningToolsJSON+26.9 intel · $0.050/M
    1.0M
    Context
  19. 19google logo
    gemini-2.5-flash
    ReasoningToolsJSON+214.1 intel · $0.300/M
    1.0M
    Context
  20. 20google logo
    gemini-2.5-flash:batch
    ReasoningToolsJSON+214.1 intel · $0.150/M
    1.0M
    Context
  21. 21google logo
    gemini-2.5-pro
    ReasoningToolsJSON+225.8 intel · $1.25/M
    1.0M
    Context
  22. 22google logo
    gemini-2.5-pro:batch
    ReasoningToolsJSON+225.8 intel · $0.625/M
    1.0M
    Context
  23. 23google logo
    gemini-2.5-pro-preview
    ReasoningToolsJSON+2$1.25/M
    1.0M
    Context
  24. 24google logo
    gemini-2.5-pro-preview-05-06
    ReasoningToolsJSON+2$1.25/M
    1.0M
    Context
  25. 25google logo
    gemma-4-26b-a4b-it
    ReasoningToolsJSON+1$0.070/M
    262K
    Context

Frequently asked

Which Google model has the largest context window?

Gemini 3.6 Flash has the largest context window of any Google model, at 1.0M tokens. Gemini 3.6 Flash (batch) (1.0M) and Gemini 3.5 Flash Lite (1.0M) round out the top three.

What's a good alternative to Gemini 3.6 Flash?

Gemini 3.6 Flash (batch) (1.0M) is the closest alternative on this metric, followed by Gemini 3.5 Flash Lite (1.0M). See the full ranking above for the tradeoffs.

How many Google models are there?

modelgrep tracks 39 Google models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Gemini 3.5 Flash. 25 of them qualify for this ranking.

More Google rankings

All rankings