modelgrep
google logo

Google: Gemini 3.6 Flash

google/gemini-3.6-flash

15th smartest of 155ReasoningToolsJSONVisionAudio
Use via OpenRouter ↗
Intelligence
50.1
15th of 155
Design Elo
Speed
175
1st fastest
Latency
1.7s
first token
Input price
$1.50
236th cheapest
Context
1.0M
66K max out

How it compares

Smarter than90%
of all ranked models
Faster than95%
of all ranked models
Cheaper than24%
of all ranked models

Overview

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Benchmarks

independent · Artificial Analysis & Design Arena
Artificial Analysis96th percentile
Intelligence Index
50.1
Coding Index
69.2
GPQA Diamond
93%
Humanity's Last Exam
38%
SciCode
53%

Providers & pricing (6)

ProviderIn $/MOut $/MUptime
Googlecache$1.50$7.5099.8%
Googlecache$0.750$3.7599.8%
Googlecache$2.70$13.5099.8%
Google AI Studiocache$1.50$7.5099.5%
Google AI Studiocache$0.750$3.7599.5%
Google AI Studiocache$2.70$13.5099.5%

Specifications

Context window1.0M
Max output66K
Knowledge cutoff
Input modalitiestext, image, video, file, audio
Output modalitiestext
Prompt cachingSupported
Cache read price$0.150/M
ModeratedNo

Gemini 3.6 Flash FAQ

How much does Gemini 3.6 Flash cost?

Gemini 3.6 Flash costs $1.50 per million input tokens and $7.50 per million output tokens via OpenRouter, making it 236th cheapest of 310 paid models.

How smart is Gemini 3.6 Flash?

Gemini 3.6 Flash scores 50.1 on the Artificial Analysis Intelligence Index, ranking 15th of 155 benchmarked models, with a GPQA Diamond score of 93%.

How fast is Gemini 3.6 Flash?

Gemini 3.6 Flash generates around 175 tokens per second with 1.7s time-to-first-token (p50), the 1st fastest tracked model.

What is Gemini 3.6 Flash's context window?

Gemini 3.6 Flash supports a 1.0M-token context window and can output up to 66K tokens. It accepts text, image, video, file, audio input.

Compare head-to-head