modelgrep

Best LLMs for Instruction Following — Meta

Match · Updated August 2026

The best Meta model for instruction following is Llama 4 Maverick, satisfying 43.0% of IFBench's verifiable prompt constraints. Llama 4 Scout (39.5%) is next.

43.0%IFBench
14.3Intelligence
60 t/sSpeed
$0.200Input /M
1.0MContext
  1. 1meta-llama logo
    llama-4-maverick
    ToolsJSONVision14.3 intel · $0.200/M · 60 t/s
    43.0%
    IFBench
  2. 2meta-llama logo
    llama-4-scout
    ToolsJSONVision10.0 intel · $0.100/M · 92 t/s
    39.5%
    IFBench

How this is ranked

AI models ranked by IFBench — how reliably a model obeys explicit, verifiable constraints in the prompt (format, length, inclusion and exclusion rules). High scores mean fewer retries and less prompt-wrangling in production, which often matters more than raw intelligence for pipelines.

Frequently asked

Which Meta model follows instructions most reliably?

The best Meta model for instruction following is Llama 4 Maverick, satisfying 43.0% of IFBench's verifiable prompt constraints. Llama 4 Scout (39.5%) is next.

What's a good alternative to Llama 4 Maverick?

Llama 4 Scout (39.5%) is the closest alternative on this metric. See the full ranking above for the tradeoffs.

How many Meta models are there?

modelgrep tracks 8 Meta models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Llama 4 Maverick. 2 of them qualify for this ranking.

More Meta rankings

All rankings