Claude Opus 5 is the best vision-capable LLM, pairing 60.7 intelligence with image and document understanding. Claude Fable 5 (59.9) and Claude Fable 5 (batch) (59.9) round out the top three.
Multimodal large language models that accept image input, ranked by intelligence. The best vision language models (VLMs) for understanding images, documents and charts.
Claude Opus 5 is the best vision-capable LLM, pairing 60.7 intelligence with image and document understanding. Claude Fable 5 (59.9) and Claude Fable 5 (batch) (59.9) round out the top three.
Claude Opus 5 is the best vision-capable AI model, pairing 60.7 intelligence with image and document understanding. Claude Fable 5 (59.9) and Claude Fable 5 (batch) (59.9) round out the top three.
Claude Fable 5 (59.9) is the closest alternative on this metric, followed by Claude Fable 5 (batch) (59.9). See the full ranking above for the tradeoffs.