modelgrep

Best LLMs for RAG

Match · Updated July 2026

The best LLM for RAG is Claude Fable 5 — 59.9 intelligence with 1M tokens of context for retrieved passages. GPT-5.6 Sol (58.9) and Kimi K3 (57.1) round out the top three.

59.9Intelligence
$10.00Input /M
1MContext

The best AI models for retrieval-augmented generation, ranked by intelligence among models with at least 128K tokens of context — enough to hold retrieved passages plus conversation. RAG pipelines run at volume, so weigh the price and speed columns as hard as the score.

  1. 1anthropic logo
    claude-fable-5
    ReasoningToolsJSON+159.9 intel · $10.00/M · 1M ctx
    59.9
    Intelligence
  2. 2openai logo
    gpt-5.6-sol
    ReasoningToolsJSON+158.9 intel · $5.00/M · 74 t/s
    58.9
    Intelligence
  3. 3moonshotai logo
    kimi-k3
    ReasoningToolsJSON+157.1 intel · $3.00/M · 1.0M ctx
    57.1
    Intelligence
  4. 4anthropic logo
    claude-opus-4.8
    ReasoningToolsJSON+155.7 intel · $5.00/M · 1M ctx
    55.7
    Intelligence
  5. 5openai logo
    gpt-5.6-terra
    ReasoningToolsJSON+155.0 intel · $2.50/M · 96 t/s
    55.0
    Intelligence
  6. 6openai logo
    gpt-5.5
    ReasoningToolsJSON+154.8 intel · $5.00/M · 100 t/s
    54.8
    Intelligence
  7. 7x-ai logo
    grok-4.5
    ReasoningToolsJSON+153.8 intel · $2.00/M · 500K ctx
    53.8
    Intelligence
  8. 8anthropic logo
    claude-opus-4.7
    ReasoningToolsJSON+153.5 intel · $5.00/M · 1M ctx
    53.5
    Intelligence
  9. 9anthropic logo
    claude-sonnet-5
    ReasoningToolsJSON+153.4 intel · $2.00/M · 1M ctx
    53.4
    Intelligence
  10. 10openai logo
    gpt-5.4
    ReasoningToolsJSON+151.4 intel · $2.50/M · 79 t/s
    51.4
    Intelligence
  11. 11openai logo
    gpt-5.6-luna
    ReasoningToolsJSON+151.2 intel · $1.00/M · 97 t/s
    51.2
    Intelligence
  12. 12z-ai logo
    glm-5.2
    ReasoningToolsJSON51.1 intel · $0.808/M · 1.0M ctx
    51.1
    Intelligence
  13. 13M
    muse-spark-1.1
    ReasoningToolsJSON+250.6 intel · $1.25/M · 1.0M ctx
    50.6
    Intelligence
  14. 14google logo
    gemini-3.5-flash
    ReasoningToolsJSON+250.2 intel · $1.50/M · 189 t/s
    50.2
    Intelligence
  15. 15google logo
    gemini-3.6-flash
    ReasoningToolsJSON+250.1 intel · $1.50/M · 111 t/s
    50.1
    Intelligence
  16. 16google logo
    gemini-3.1-pro-preview
    ReasoningToolsJSON+246.5 intel · $2.00/M · 122 t/s
    46.5
    Intelligence
  17. 17qwen logo
    qwen3.7-max
    ReasoningToolsJSON46.0 intel · $1.48/M · 1M ctx
    46.0
    Intelligence
  18. 18minimax logo
    minimax-m3
    ReasoningToolsJSON+144.4 intel · $0.300/M · 1.0M ctx
    44.4
    Intelligence
  19. 19deepseek logo
    deepseek-v4-pro
    ReasoningToolsJSON44.3 intel · $0.435/M · 1.0M ctx
    44.3
    Intelligence
  20. 20openai logo
    gpt-5.3-codex
    ReasoningToolsJSON+144.3 intel · $1.75/M · 400K ctx
    44.3
    Intelligence
  21. 21moonshotai logo
    kimi-k2.6
    ReasoningToolsJSON+144.2 intel · $0.684/M · 262K ctx
    44.2
    Intelligence
  22. 22xiaomi logo
    mimo-v2.5-pro
    ReasoningToolsJSON42.2 intel · $0.435/M · 1.1M ctx
    42.2
    Intelligence
  23. 23openai logo
    gpt-5.2
    ReasoningToolsJSON+142.2 intel · $1.75/M · 71 t/s
    42.2
    Intelligence
  24. 24moonshotai logo
    kimi-k2.7-code
    ReasoningToolsJSON+141.9 intel · $0.710/M · 262K ctx
    41.9
    Intelligence
  25. 25tencent logo
    hy3
    ReasoningToolsJSON41.2 intel · $0.132/M · 262K ctx
    41.2
    Intelligence

Frequently asked

What is the best LLM for RAG?

The best LLM for RAG is Claude Fable 5 — 59.9 intelligence with 1M tokens of context for retrieved passages. GPT-5.6 Sol (58.9) and Kimi K3 (57.1) round out the top three.

Which AI model works best in a RAG pipeline?

The best AI model for RAG is Claude Fable 5 — 59.9 intelligence with 1M tokens of context for retrieved passages. GPT-5.6 Sol (58.9) and Kimi K3 (57.1) round out the top three.

What's a good alternative to Claude Fable 5?

GPT-5.6 Sol (58.9) is the closest alternative on this metric, followed by Kimi K3 (57.1). See the full ranking above for the tradeoffs.

By maker

All rankings