modelgrep
nvidia logo

NVIDIA: Nemotron 3 Super

nvidia/nemotron-3-super-120b-a12b

Cheaper than 88% of paidReasoningToolsJSON
Use via OpenRouter ↗
Intelligence
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.085
39th cheapest
Context
1M
16K max out

How it compares

Cheaper than88%
of all ranked models

Overview

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

Providers & pricing (3)

ProviderIn $/MOut $/MUptime
DeepInfrabf16$0.085$0.40088.4%
DigitalOcean$0.210$0.45598%
Nebiusfp4$0.300$0.900100%

Specifications

Context window1M
Max output16K
Knowledge cutoff
Input modalitiestext
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Nemotron 3 Super FAQ

How much does Nemotron 3 Super cost?

Nemotron 3 Super costs $0.085 per million input tokens and $0.400 per million output tokens via OpenRouter, making it 39th cheapest of 332 paid models.

What is Nemotron 3 Super's context window?

Nemotron 3 Super supports a 1M-token context window and can output up to 16K tokens. It accepts text input.

Compare head-to-head