modelgrep

Best LLMs for Long-Context Reasoning — OpenAI

Match · Updated August 2026

The best OpenAI model for reasoning over long inputs is GPT-5.2-Codex, scoring 75.7% on long-context reasoning — a different question from how large a window it accepts. GPT-5 (75.6%) and GPT-5.1 (75.0%) round out the top three.

75.7%LCR
40.1Intelligence
$1.75Input /M
400KContext
  1. 1openai logo
    gpt-5.2-codex
    ReasoningToolsJSON+140.1 intel · $1.75/M · 400K ctx
    75.7%
    LCR
  2. 2openai logo
    gpt-5
    ReasoningToolsJSON+134.7 intel · $1.25/M · 400K ctx
    75.6%
    LCR
  3. 3openai logo
    gpt-5.1
    ReasoningToolsJSON+136.9 intel · $1.25/M · 400K ctx
    75.0%
    LCR
  4. 4openai logo
    gpt-5.5
    ReasoningToolsJSON+154.8 intel · $5.00/M · 72 t/s
    74.3%
    LCR
  5. 5openai logo
    gpt-5.6-luna
    ReasoningToolsJSON+151.2 intel · $0.100/M · 86 t/s
    74.0%
    LCR
  6. 6openai logo
    gpt-5.6-terra
    ReasoningToolsJSON+155.0 intel · $1.00/M · 60 t/s
    74.0%
    LCR
  7. 7openai logo
    gpt-5.4
    ReasoningToolsJSON+151.4 intel · $2.50/M · 72 t/s
    74.0%
    LCR
  8. 8openai logo
    gpt-5.3-codex
    ReasoningToolsJSON+144.3 intel · $1.75/M · 55 t/s
    74.0%
    LCR
  9. 9openai logo
    gpt-5.6-sol
    ReasoningToolsJSON+158.9 intel · $5.00/M · 49 t/s
    73.7%
    LCR
  10. 10openai logo
    gpt-5.2
    ReasoningToolsJSON+142.2 intel · $1.75/M · 400K ctx
    72.7%
    LCR
  11. 11openai logo
    gpt-5.4-mini
    ReasoningToolsJSON+140.0 intel · $0.750/M · 132 t/s
    69.3%
    LCR
  12. 12openai logo
    o3
    ReasoningToolsJSON+130.4 intel · $2.00/M · 44 t/s
    69.3%
    LCR
  13. 13openai logo
    gpt-5-mini
    ReasoningToolsJSON+125.3 intel · $0.250/M · 400K ctx
    68.0%
    LCR
  14. 14openai logo
    gpt-5.1-codex
    ReasoningToolsJSON+134.7 intel · $1.25/M · 400K ctx
    67.3%
    LCR
  15. 15openai logo
    gpt-5.4-nano
    ReasoningToolsJSON+138.2 intel · $0.200/M · 45 t/s
    66.0%
    LCR
  16. 16openai logo
    gpt-5.1-codex-mini
    ReasoningToolsJSON+130.6 intel · $0.250/M · 400K ctx
    62.7%
    LCR
  17. 17openai logo
    gpt-4.1
    ToolsJSONVision19.4 intel · $2.00/M · 62 t/s
    61.0%
    LCR
  18. 18openai logo
    o1
    ReasoningToolsJSON+123.4 intel · $15.00/M · 48 t/s
    59.3%
    LCR
  19. 19openai logo
    o4-mini-high
    ReasoningToolsJSON+125.6 intel · $1.10/M · 95 t/s
    55.0%
    LCR
  20. 20openai logo
    o4-mini
    ReasoningToolsJSON+125.6 intel · $1.10/M · 81 t/s
    55.0%
    LCR
  21. 21openai logo
    gpt-oss-120b
    ReasoningToolsJSON23.8 intel · $0.037/M · 199 t/s
    50.7%
    LCR
  22. 22openai logo
    gpt-4.1-mini
    ToolsJSONVision14.8 intel · $0.400/M · 43 t/s
    42.3%
    LCR
  23. 23openai logo
    gpt-5-nano
    ReasoningToolsJSON+119.9 intel · $0.050/M · 400K ctx
    41.7%
    LCR
  24. 24openai logo
    o3-mini-high
    ReasoningToolsJSON15.6 intel · $1.10/M · 115 t/s
    39.3%
    LCR
  25. 25openai logo
    gpt-4o-2024-08-06
    ToolsJSONVision9.6 intel · $2.50/M · 28 t/s
    35.0%
    LCR

How this is ranked

AI models ranked by LCR — reasoning accuracy over long inputs, not just the size of the window they accept. A large context window is a capacity claim; this is the measurement of whether the model can still reason over information buried deep inside it. Pair it with the longest-context ranking, which sorts on raw window size.

Frequently asked

Which OpenAI model reasons best over long inputs?

The best OpenAI model for reasoning over long inputs is GPT-5.2-Codex, scoring 75.7% on long-context reasoning — a different question from how large a window it accepts. GPT-5 (75.6%) and GPT-5.1 (75.0%) round out the top three.

What's a good alternative to GPT-5.2-Codex?

GPT-5 (75.6%) is the closest alternative on this metric, followed by GPT-5.1 (75.0%). See the full ranking above for the tradeoffs.

How many OpenAI models are there?

modelgrep tracks 60 OpenAI models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by GPT-5.6 Sol. 25 of them qualify for this ranking.

More OpenAI rankings

All rankings