modelgrep

Best Local MiniMax Models

Match · Updated July 2026

The best local MiniMax model is MiniMax M3 (44.4 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. MiniMax M2.7 (38.1) and MiniMax M2.5 (33.7) round out the top three.

44.4Intelligence
$0.300Input /M
1.0MContext

The best local LLMs — open-weight models small enough to run on your own hardware with Ollama, llama.cpp or vLLM — ranked by intelligence. Frontier-scale open models (400B+) are excluded: downloadable isn't the same as runnable. Every model here is on Hugging Face.

  1. 1minimax logo
    minimax-m3
    ReasoningToolsJSON+144.4 intel · $0.300/M · 1.0M ctx
    44.4
    Intelligence
  2. 2minimax logo
    minimax-m2.7
    ReasoningToolsJSON38.1 intel · $0.250/M · 205K ctx
    38.1
    Intelligence
  3. 3minimax logo
    minimax-m2.5
    ReasoningToolsJSON33.7 intel · $0.150/M · 205K ctx
    33.7
    Intelligence
  4. 4minimax logo
    minimax-m2.1
    ReasoningToolsJSON31.4 intel · $0.300/M · 205K ctx
    31.4
    Intelligence
  5. 5minimax logo
    minimax-m2
    ReasoningToolsJSON28.3 intel · $0.255/M · 205K ctx
    28.3
    Intelligence

Frequently asked

What is the best local MiniMax model?

The best local MiniMax model is MiniMax M3 (44.4 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. MiniMax M2.7 (38.1) and MiniMax M2.5 (33.7) round out the top three.

What's a good alternative to MiniMax M3?

MiniMax M2.7 (38.1) is the closest alternative on this metric, followed by MiniMax M2.5 (33.7). See the full ranking above for the tradeoffs.

How many MiniMax models are there?

modelgrep tracks 8 MiniMax models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by MiniMax M3. 5 of them qualify for this ranking.

More MiniMax rankings

All rankings