The best local Cohere model is Command A (22.5 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. North Mini Code (free) (19.8) is next.
The best local LLMs — open-weight models small enough to run on your own hardware with Ollama, llama.cpp or vLLM — ranked by intelligence. Frontier-scale open models (400B+) are excluded: downloadable isn't the same as runnable. Every model here is on Hugging Face.
The best local Cohere model is Command A (22.5 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. North Mini Code (free) (19.8) is next.
North Mini Code (free) (19.8) is the closest alternative on this metric. See the full ranking above for the tradeoffs.
modelgrep tracks 5 Cohere models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Command A. 2 of them qualify for this ranking.