The best local Meta model is Llama 4 Maverick (14.3 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. Llama 4 Scout (10.0) is next.
The best local LLMs — open-weight models small enough to run on your own hardware with Ollama, llama.cpp or vLLM — ranked by intelligence. Frontier-scale open models (400B+) are excluded: downloadable isn't the same as runnable. Every model here is on Hugging Face.
The best local Meta model is Llama 4 Maverick (14.3 intelligence) — open weights on Hugging Face, small enough to run on your own hardware. Llama 4 Scout (10.0) is next.
Llama 4 Scout (10.0) is the closest alternative on this metric. See the full ranking above for the tradeoffs.
modelgrep tracks 8 Meta models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Llama 4 Maverick. 2 of them qualify for this ranking.