The best Qwen model for reasoning over long inputs is Qwen3.6 Max Preview, scoring 69.7% on long-context reasoning — a different question from how large a window it accepts. Qwen3.6 Plus (69.7%) and Qwen3.7 Max (69.0%) round out the top three.
AI models ranked by LCR — reasoning accuracy over long inputs, not just the size of the window they accept. A large context window is a capacity claim; this is the measurement of whether the model can still reason over information buried deep inside it. Pair it with the longest-context ranking, which sorts on raw window size.
The best Qwen model for reasoning over long inputs is Qwen3.6 Max Preview, scoring 69.7% on long-context reasoning — a different question from how large a window it accepts. Qwen3.6 Plus (69.7%) and Qwen3.7 Max (69.0%) round out the top three.
Qwen3.6 Plus (69.7%) is the closest alternative on this metric, followed by Qwen3.7 Max (69.0%). See the full ranking above for the tradeoffs.
modelgrep tracks 49 Qwen models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Qwen3.7 Max. 21 of them qualify for this ranking.