modelgrep

Best Local MoonshotAI Models

Match · Updated July 2026

The best local MoonshotAI model is Kimi K2.6 (44.2 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. Kimi K2.7 Code (41.9) and Kimi K2.5 (35.4) round out the top three.

44.2Intelligence
$0.684Input /M
262KContext

The best local LLMs — open-weight models small enough to run on your own hardware with Ollama, llama.cpp or vLLM — ranked by intelligence. Frontier-scale open models (400B+) are excluded: downloadable isn't the same as runnable. Every model here is on Hugging Face.

  1. 1moonshotai logo
    kimi-k2.6
    ReasoningToolsJSON+144.2 intel · $0.684/M · 262K ctx
    44.2
    Intelligence
  2. 2moonshotai logo
    kimi-k2.7-code
    ReasoningToolsJSON+141.9 intel · $0.710/M · 262K ctx
    41.9
    Intelligence
  3. 3moonshotai logo
    kimi-k2.5
    ReasoningToolsJSON+135.4 intel · $0.570/M · 262K ctx
    35.4
    Intelligence
  4. 4moonshotai logo
    kimi-k2-thinking
    ReasoningToolsJSON32.7 intel · $0.600/M · 262K ctx
    32.7
    Intelligence
  5. 5moonshotai logo
    kimi-k2-0905
    ToolsJSON23.5 intel · $0.600/M · 262K ctx
    23.5
    Intelligence
  6. 6moonshotai logo
    kimi-k2
    Tools19.4 intel · $0.570/M · 131K ctx
    19.4
    Intelligence

Frequently asked

What is the best local MoonshotAI model?

The best local MoonshotAI model is Kimi K2.6 (44.2 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. Kimi K2.7 Code (41.9) and Kimi K2.5 (35.4) round out the top three.

What's a good alternative to Kimi K2.6?

Kimi K2.7 Code (41.9) is the closest alternative on this metric, followed by Kimi K2.5 (35.4). See the full ranking above for the tradeoffs.

How many MoonshotAI models are there?

modelgrep tracks 7 MoonshotAI models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Kimi K3. 6 of them qualify for this ranking.

More MoonshotAI rankings

All rankings