modelgrep

Best LLMs for Long-Context Reasoning — Meta

Match · Updated August 2026

The best Meta model for reasoning over long inputs is Llama 4 Maverick, scoring 46.0% on long-context reasoning — a different question from how large a window it accepts. Llama 4 Scout (25.8%) is next.

46.0%LCR
14.3Intelligence
107 t/sSpeed
$0.200Input /M
1.0MContext
  1. 1meta-llama logo
    llama-4-maverick
    ToolsJSONVision14.3 intel · $0.200/M · 107 t/s
    46.0%
    LCR
  2. 2meta-llama logo
    llama-4-scout
    ToolsJSONVision10.0 intel · $0.100/M · 121 t/s
    25.8%
    LCR

How this is ranked

AI models ranked by LCR — reasoning accuracy over long inputs, not just the size of the window they accept. A large context window is a capacity claim; this is the measurement of whether the model can still reason over information buried deep inside it. Pair it with the longest-context ranking, which sorts on raw window size.

Frequently asked

Which Meta model reasons best over long inputs?

The best Meta model for reasoning over long inputs is Llama 4 Maverick, scoring 46.0% on long-context reasoning — a different question from how large a window it accepts. Llama 4 Scout (25.8%) is next.

What's a good alternative to Llama 4 Maverick?

Llama 4 Scout (25.8%) is the closest alternative on this metric. See the full ranking above for the tradeoffs.

How many Meta models are there?

modelgrep tracks 8 Meta models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Llama 4 Maverick. 2 of them qualify for this ranking.

More Meta rankings

All rankings