modelgrep
qwen logo

Qwen: Qwen3 Next 80B A3B Thinking

qwen/qwen3-next-80b-a3b-thinking

ReasoningToolsJSON
Use via OpenRouter ↗
Intelligence
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.150
86th cheapest
Context
262K
33K max out

How it compares

Cheaper than74%
of all ranked models

Overview

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

Providers & pricing (3)

ProviderIn $/MOut $/MUptime
Google$0.150$1.20
Nebiusfp8$0.150$1.20
Alibabafp8$0.150$1.20

Specifications

Context window262K
Max output33K
Knowledge cutoffSep 2025
Input modalitiestext
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

Qwen3 Next 80B A3B Thinking FAQ

How much does Qwen3 Next 80B A3B Thinking cost?

Qwen3 Next 80B A3B Thinking costs $0.150 per million input tokens and $1.20 per million output tokens via OpenRouter, making it 86th cheapest of 332 paid models.

What is Qwen3 Next 80B A3B Thinking's context window?

Qwen3 Next 80B A3B Thinking supports a 262K-token context window and can output up to 33K tokens. It accepts text input.

Compare head-to-head