modelgrep

Best LLMs for Long-Context Reasoning — Qwen

Match · Updated August 2026

The best Qwen model for reasoning over long inputs is Qwen3.6 Max Preview, scoring 69.7% on long-context reasoning — a different question from how large a window it accepts. Qwen3.6 Plus (69.7%) and Qwen3.7 Max (69.0%) round out the top three.

69.7%LCR
40.0Intelligence
$1.03Input /M
262KContext
  1. 1qwen logo
    qwen3.6-max-preview
    ReasoningToolsJSON40.0 intel · $1.03/M · 262K ctx
    69.7%
    LCR
  2. 2qwen logo
    qwen3.6-plus
    ReasoningToolsJSON+139.6 intel · $0.325/M · 55 t/s
    69.7%
    LCR
  3. 3qwen logo
    qwen3.7-max
    ReasoningToolsJSON46.0 intel · $1.48/M · 207 t/s
    69.0%
    LCR
  4. 4qwen logo
    qwen3.6-27b
    ReasoningToolsJSON+137.1 intel · $0.289/M · 60 t/s
    68.7%
    LCR
  5. 5qwen logo
    qwen3.5-27b
    ReasoningToolsJSON+133.8 intel · $0.195/M · 262K ctx
    67.3%
    LCR
  6. 6qwen logo
    qwen3.5-122b-a10b
    ReasoningToolsJSON+132.3 intel · $0.260/M · 127 t/s
    66.7%
    LCR
  7. 7qwen logo
    qwen3-max-thinking
    ReasoningToolsJSON31.7 intel · $0.780/M · 262K ctx
    66.0%
    LCR
  8. 8qwen logo
    qwen3.5-397b-a17b
    ReasoningToolsJSON+133.7 intel · $0.390/M · 69 t/s
    65.7%
    LCR
  9. 9qwen logo
    qwen3.7-plus
    ReasoningToolsJSON+139.0 intel · $0.320/M · 54 t/s
    65.0%
    LCR
  10. 10qwen logo
    qwen3.6-35b-a3b
    ReasoningToolsJSON+131.6 intel · $0.140/M · 158 t/s
    63.7%
    LCR
  11. 11qwen logo
    qwen3.5-35b-a3b
    ReasoningToolsJSON+129.3 intel · $0.140/M · 262K ctx
    62.7%
    LCR
  12. 12qwen logo
    qwen3.5-9b
    ReasoningToolsJSON+121.4 intel · $0.100/M · 74 t/s
    59.0%
    LCR
  13. 13qwen logo
    qwen3-next-80b-a3b-instruct
    ToolsJSON13.7 intel · $0.090/M · 205 t/s
    51.3%
    LCR
  14. 14qwen logo
    qwen3-max
    ToolsJSON24.0 intel · $0.780/M · 262K ctx
    46.7%
    LCR
  15. 15qwen logo
    qwen3-coder-next
    ToolsJSON21.1 intel · $0.120/M · 72 t/s
    40.0%
    LCR
  16. 16qwen logo
    qwen3-vl-235b-a22b-instruct
    ToolsJSONVision14.3 intel · $0.210/M · 262K ctx
    31.7%
    LCR
  17. 17qwen logo
    qwen3-vl-32b-instruct
    ToolsJSONVision11.1 intel · $0.104/M · 38 t/s
    31.3%
    LCR
  18. 18qwen logo
    qwen3-coder-30b-a3b-instruct
    ToolsJSON13.6 intel · $0.070/M · 262K ctx
    29.0%
    LCR
  19. 19qwen logo
    qwen3-vl-30b-a3b-instruct
    ToolsJSONVision10.0 intel · $0.150/M · 34 t/s
    23.7%
    LCR
  20. 20qwen logo
    qwen-2.5-72b-instruct
    ToolsJSON9.6 intel · $0.360/M · 24 t/s
    20.3%
    LCR
  21. 21qwen logo
    qwen3-vl-8b-instruct
    ToolsJSONVision8.4 intel · $0.117/M · 42 t/s
    15.3%
    LCR

How this is ranked

AI models ranked by LCR — reasoning accuracy over long inputs, not just the size of the window they accept. A large context window is a capacity claim; this is the measurement of whether the model can still reason over information buried deep inside it. Pair it with the longest-context ranking, which sorts on raw window size.

Frequently asked

Which Qwen model reasons best over long inputs?

The best Qwen model for reasoning over long inputs is Qwen3.6 Max Preview, scoring 69.7% on long-context reasoning — a different question from how large a window it accepts. Qwen3.6 Plus (69.7%) and Qwen3.7 Max (69.0%) round out the top three.

What's a good alternative to Qwen3.6 Max Preview?

Qwen3.6 Plus (69.7%) is the closest alternative on this metric, followed by Qwen3.7 Max (69.0%). See the full ranking above for the tradeoffs.

How many Qwen models are there?

modelgrep tracks 49 Qwen models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Qwen3.7 Max. 21 of them qualify for this ranking.

More Qwen rankings

All rankings