modelgrep
qwen logo

Qwen: Qwen3 235B A22B Instruct 2507

qwen/qwen3-235b-a22b-2507

Cheaper than 88% of paidToolsJSON
Use via OpenRouter ↗
Intelligence
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.090
40th cheapest
Context
262K
16K max out

How it compares

Cheaper than88%
of all ranked models

Overview

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

Providers & pricing (12)

ProviderIn $/MOut $/MUptime
DeepInfrafp8$0.090$0.55099.5%
Novitafp8$0.090$0.58099.3%
Parasailfp8$0.140$0.80099.5%
Alibabafp8$0.149$0.598100%
Venicefp8$0.150$0.75099.6%
Nebiusfp8$0.200$0.60090%
Friendli$0.200$0.80097.7%
AtlasCloudfp8$0.200$0.88099.8%
StreamLake$0.210$0.840100%
Crusoebf16$0.220$0.80099%
Google$0.220$0.880100%
Google$0.250$1.0099.2%

Specifications

Context window262K
Max output16K
Knowledge cutoffJun 2025
Input modalitiestext
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Qwen3 235B A22B Instruct 2507 FAQ

How much does Qwen3 235B A22B Instruct 2507 cost?

Qwen3 235B A22B Instruct 2507 costs $0.090 per million input tokens and $0.550 per million output tokens via OpenRouter, making it 40th cheapest of 332 paid models.

What is Qwen3 235B A22B Instruct 2507's context window?

Qwen3 235B A22B Instruct 2507 supports a 262K-token context window and can output up to 16K tokens. It accepts text input.

Compare head-to-head