modelgrep

Best LLMs for Long-Context Reasoning — Mistral

Match · Updated August 2026

The best Mistral model for reasoning over long inputs is Mistral Medium 3.5, scoring 61.0% on long-context reasoning — a different question from how large a window it accepts. Mistral Medium 3 (28.0%) and Mistral Medium 3.1 (19.7%) round out the top three.

61.0%LCR
29.9Intelligence
152 t/sSpeed
$1.50Input /M
262KContext
  1. 1mistralai logo
    mistral-medium-3-5
    ReasoningToolsJSON+129.9 intel · $1.50/M · 152 t/s
    61.0%
    LCR
  2. 2mistralai logo
    mistral-medium-3
    ToolsJSONVision12.5 intel · $0.400/M · 131K ctx
    28.0%
    LCR
  3. 3mistralai logo
    mistral-medium-3.1
    ToolsJSONVision14.7 intel · $0.400/M · 131K ctx
    19.7%
    LCR
  4. 4mistralai logo
    mistral-large-2407
    ToolsJSON7.3 intel · $2.00/M · 26 t/s
    1.7%
    LCR

How this is ranked

AI models ranked by LCR — reasoning accuracy over long inputs, not just the size of the window they accept. A large context window is a capacity claim; this is the measurement of whether the model can still reason over information buried deep inside it. Pair it with the longest-context ranking, which sorts on raw window size.

Frequently asked

Which Mistral model reasons best over long inputs?

The best Mistral model for reasoning over long inputs is Mistral Medium 3.5, scoring 61.0% on long-context reasoning — a different question from how large a window it accepts. Mistral Medium 3 (28.0%) and Mistral Medium 3.1 (19.7%) round out the top three.

What's a good alternative to Mistral Medium 3.5?

Mistral Medium 3 (28.0%) is the closest alternative on this metric, followed by Mistral Medium 3.1 (19.7%). See the full ranking above for the tradeoffs.

How many Mistral models are there?

modelgrep tracks 18 Mistral models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Mistral Medium 3.5. 4 of them qualify for this ranking.

More Mistral rankings

All rankings