modelgrep
google logo

Google: Gemma 3 4B

google/gemma-3-4b-it

Cheaper than 95% of paidJSONVision
Use via OpenRouter ↗
Intelligence
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.050
18th cheapest
Context
131K
16K max out

How it compares

Cheaper than95%
of all ranked models

Overview

Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...

Providers & pricing (1)

ProviderIn $/MOut $/MUptime
DeepInfrabf16$0.050$0.100100%

Specifications

Context window131K
Max output16K
Knowledge cutoffAug 2024
Input modalitiestext, image
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Gemma 3 4B FAQ

How much does Gemma 3 4B cost?

Gemma 3 4B costs $0.050 per million input tokens and $0.100 per million output tokens via OpenRouter, making it 18th cheapest of 332 paid models.

What is Gemma 3 4B's context window?

Gemma 3 4B supports a 131K-token context window and can output up to 16K tokens. It accepts text, image input.

Compare head-to-head