modelgrep
qwen logo

Qwen: Qwen3 30B A3B Instruct 2507

qwen/qwen3-30b-a3b-instruct-2507

Cheaper than 96% of paidToolsJSON
Use via OpenRouter ↗
Intelligence
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.048
13th cheapest
Context
262K
32K max out

How it compares

Cheaper than96%
of all ranked models

Overview

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...

Providers & pricing (5)

ProviderIn $/MOut $/MUptime
StreamLake$0.048$0.193100%
SiliconFlowfp8$0.090$0.30096.9%
Nebiusfp8$0.100$0.30098.6%
CoreWeavebf16$0.100$0.300100%
Alibabafp8$0.130$0.520100%

Specifications

Context window262K
Max output32K
Knowledge cutoffJun 2025
Input modalitiestext
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Qwen3 30B A3B Instruct 2507 FAQ

How much does Qwen3 30B A3B Instruct 2507 cost?

Qwen3 30B A3B Instruct 2507 costs $0.048 per million input tokens and $0.193 per million output tokens via OpenRouter, making it 13th cheapest of 332 paid models.

What is Qwen3 30B A3B Instruct 2507's context window?

Qwen3 30B A3B Instruct 2507 supports a 262K-token context window and can output up to 32K tokens. It accepts text input.

Compare head-to-head