modelgrep

Best Local Qwen Models

Match · Updated July 2026

The best local Qwen model is Qwen3.6 27B (37.1 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. Qwen3.5-27B (33.8) and Qwen3.5 397B A17B (33.7) round out the top three.

37.1Intelligence
$0.288Input /M
262KContext

The best local LLMs — open-weight models small enough to run on your own hardware with Ollama, llama.cpp or vLLM — ranked by intelligence. Frontier-scale open models (400B+) are excluded: downloadable isn't the same as runnable. Every model here is on Hugging Face.

  1. 1qwen logo
    qwen3.6-27b
    ReasoningToolsJSON+137.1 intel · $0.288/M · 262K ctx
    37.1
    Intelligence
  2. 2qwen logo
    qwen3.5-27b
    ReasoningToolsJSON+133.8 intel · $0.195/M · 262K ctx
    33.8
    Intelligence
  3. 3qwen logo
    qwen3.5-397b-a17b
    ReasoningToolsJSON+133.7 intel · $0.390/M · 262K ctx
    33.7
    Intelligence
  4. 4qwen logo
    qwen3.5-122b-a10b
    ReasoningToolsJSON+132.3 intel · $0.260/M · 262K ctx
    32.3
    Intelligence
  5. 5qwen logo
    qwen3.6-35b-a3b
    ReasoningToolsJSON+131.6 intel · $0.140/M · 262K ctx
    31.6
    Intelligence
  6. 6qwen logo
    qwen3.5-35b-a3b
    ReasoningToolsJSON+129.3 intel · $0.140/M · 262K ctx
    29.3
    Intelligence
  7. 7qwen logo
    qwen3.5-9b
    ReasoningToolsJSON+121.4 intel · $0.100/M · 262K ctx
    21.4
    Intelligence
  8. 8qwen logo
    qwen3-coder-next
    ToolsJSON21.1 intel · $0.110/M · 262K ctx
    21.1
    Intelligence
  9. 9qwen logo
    qwen3-next-80b-a3b-instruct
    ToolsJSON13.7 intel · $0.150/M · 262K ctx
    13.7
    Intelligence
  10. 10qwen logo
    qwen3-coder-30b-a3b-instruct
    ToolsJSON13.6 intel · $0.070/M · 262K ctx
    13.6
    Intelligence
  11. 11qwen logo
    qwen3-vl-32b-instruct
    ToolsJSONVision11.1 intel · $0.104/M · 131K ctx
    11.1
    Intelligence
  12. 12qwen logo
    qwen3-vl-30b-a3b-instruct
    ToolsJSONVision10.0 intel · $0.150/M · 262K ctx
    10.0
    Intelligence
  13. 13qwen logo
    qwen3-vl-8b-instruct
    ToolsJSONVision8.4 intel · $0.117/M · 262K ctx
    8.4
    Intelligence
  14. 14qwen logo
    qwen-2.5-coder-32b-instruct
    7.1 intel · $0.660/M · 33K ctx
    7.1
    Intelligence

Frequently asked

What is the best local Qwen model?

The best local Qwen model is Qwen3.6 27B (37.1 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. Qwen3.5-27B (33.8) and Qwen3.5 397B A17B (33.7) round out the top three.

What's a good alternative to Qwen3.6 27B?

Qwen3.5-27B (33.8) is the closest alternative on this metric, followed by Qwen3.5 397B A17B (33.7). See the full ranking above for the tradeoffs.

How many Qwen models are there?

modelgrep tracks 47 Qwen models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Qwen3.7 Max. 14 of them qualify for this ranking.

More Qwen rankings

All rankings