modelgrep

Best MoonshotAI Models for Math & Science

Match · Updated July 2026

The best MoonshotAI model for math and science is Kimi K3, scoring 93.5% on GPQA Diamond (graduate-level physics, chemistry and biology). Kimi K2.6 (91.1%) and Kimi K2.7 Code (89.6%) round out the top three.

93.5%GPQA
57.1Intelligence
$3.00Input /M
1.0MContext

AI models ranked by GPQA Diamond — graduate-level physics, chemistry and biology questions that can't be answered by lookup. The best LLMs for math, quantitative reasoning and hard-science work.

  1. 1moonshotai logo
    kimi-k3
    ReasoningToolsJSON+157.1 intel · $3.00/M · 1.0M ctx
    93.5%
    GPQA
  2. 2moonshotai logo
    kimi-k2.6
    ReasoningToolsJSON+144.2 intel · $0.684/M · 262K ctx
    91.1%
    GPQA
  3. 3moonshotai logo
    kimi-k2.7-code
    ReasoningToolsJSON+141.9 intel · $0.710/M · 262K ctx
    89.6%
    GPQA
  4. 4moonshotai logo
    kimi-k2.5
    ReasoningToolsJSON+135.4 intel · $0.570/M · 262K ctx
    87.9%
    GPQA
  5. 5moonshotai logo
    kimi-k2-thinking
    ReasoningToolsJSON32.7 intel · $0.600/M · 262K ctx
    83.8%
    GPQA
  6. 6moonshotai logo
    kimi-k2-0905
    ToolsJSON23.5 intel · $0.600/M · 262K ctx
    76.7%
    GPQA
  7. 7moonshotai logo
    kimi-k2
    Tools19.4 intel · $0.570/M · 131K ctx
    76.6%
    GPQA

Frequently asked

What is the best MoonshotAI model for math?

The best MoonshotAI model for math and science is Kimi K3, scoring 93.5% on GPQA Diamond (graduate-level physics, chemistry and biology). Kimi K2.6 (91.1%) and Kimi K2.7 Code (89.6%) round out the top three.

What's a good alternative to Kimi K3?

Kimi K2.6 (91.1%) is the closest alternative on this metric, followed by Kimi K2.7 Code (89.6%). See the full ranking above for the tradeoffs.

How many MoonshotAI models are there?

modelgrep tracks 7 MoonshotAI models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Kimi K3. 7 of them qualify for this ranking.

More MoonshotAI rankings

All rankings