The best local MiniMax model is MiniMax M3 (44.4 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. MiniMax M2.7 (38.1) and MiniMax M2.5 (33.7) round out the top three.
The best local LLMs — open-weight models small enough to run on your own hardware with Ollama, llama.cpp or vLLM — ranked by intelligence. Frontier-scale open models (400B+) are excluded: downloadable isn't the same as runnable. Every model here is on Hugging Face.
The best local MiniMax model is MiniMax M3 (44.4 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. MiniMax M2.7 (38.1) and MiniMax M2.5 (33.7) round out the top three.
MiniMax M2.7 (38.1) is the closest alternative on this metric, followed by MiniMax M2.5 (33.7). See the full ranking above for the tradeoffs.
modelgrep tracks 8 MiniMax models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by MiniMax M3. 5 of them qualify for this ranking.