Claude Opus 5 is the best vision-capable Anthropic model, pairing 60.7 intelligence with image and document understanding. Claude Fable 5 (59.9) and Claude Fable 5 (batch) (59.9) round out the top three.
Multimodal large language models that accept image input, ranked by intelligence. The best vision language models (VLMs) for understanding images, documents and charts.
Claude Opus 5 is the best vision-capable Anthropic model, pairing 60.7 intelligence with image and document understanding. Claude Fable 5 (59.9) and Claude Fable 5 (batch) (59.9) round out the top three.
Claude Fable 5 (59.9) is the closest alternative on this metric, followed by Claude Fable 5 (batch) (59.9). See the full ranking above for the tradeoffs.
modelgrep tracks 26 Anthropic models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Claude Opus 5. 25 of them qualify for this ranking.