modelgrep
google logo

Google: Gemini 3.5 Flash

google/gemini-3.5-flash

21st smartest of 182ReasoningToolsJSONVisionAudio
Use via OpenRouter ↗
Intelligence
50.2
21st of 182
Design Elo
1295
svg
Speed
tokens/sec
Latency
first token
Input price
$1.50
256th cheapest
Context
1.0M
66K max out

How it compares

Smarter than88%
of all ranked models
Cheaper than23%
of all ranked models

Overview

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Benchmarks

independent · Artificial Analysis & Design Arena
Artificial Analysis96th percentile
Intelligence Index
50.2
Coding Index
70.1
GPQA Diamond
92%
Humanity's Last Exam
41%
SciCode
53%
Tau²-Bench (agentic)
95%
Design Arena · Elo2,192 tournaments
svg
1295
3D
1295
Website
1282
Data Viz
1259

Providers & pricing (6)

ProviderIn $/MOut $/MUptime
Googlecache$1.50$9.0099.8%
Googlecache$0.750$4.5099.8%
Googlecache$2.70$16.2099.8%
Google AI Studiocache$1.50$9.00100%
Google AI Studiocache$0.750$4.50100%
Google AI Studiocache$2.70$16.20100%

Specifications

Context window1.0M
Max output66K
Knowledge cutoffJan 2025
Input modalitiestext, image, video, file, audio
Output modalitiestext
Prompt cachingSupported
Cache read price$0.150/M
ModeratedNo

Gemini 3.5 Flash FAQ

How much does Gemini 3.5 Flash cost?

Gemini 3.5 Flash costs $1.50 per million input tokens and $9.00 per million output tokens via OpenRouter, making it 256th cheapest of 332 paid models.

How smart is Gemini 3.5 Flash?

Gemini 3.5 Flash scores 50.2 on the Artificial Analysis Intelligence Index, ranking 21st of 182 benchmarked models, with a GPQA Diamond score of 92%.

What is Gemini 3.5 Flash's context window?

Gemini 3.5 Flash supports a 1.0M-token context window and can output up to 66K tokens. It accepts text, image, video, file, audio input.

Compare head-to-head