modelgrep

Best LLMs for Tool Calling — Cohere

Match · Updated August 2026

The best Cohere model for tool calling is North Mini Code (free), completing 37.4% of Tau²-Bench's multi-turn tool-use tasks. Command A (15.2%) is next.

37.4%τ²-Bench
19.8Intelligence
18 t/sSpeed
FreeInput /M
256KContext
  1. 1cohere logo
    north-mini-code:free
    ReasoningTools19.8 intel · Free/M · 18 t/s
    37.4%
    τ²-Bench
  2. 2cohere logo
    command-a
    JSON7.7 intel · $2.50/M · 47 t/s
    15.2%
    τ²-Bench

How this is ranked

AI models ranked by Tau²-Bench — multi-turn conversations where the model has to call the right tools, in the right order, against a real API to complete a customer task. This measures whether function calling actually works under pressure, which is a different question from whether a model supports the parameter at all.

Frequently asked

Which Cohere model is best at tool calling?

The best Cohere model for tool calling is North Mini Code (free), completing 37.4% of Tau²-Bench's multi-turn tool-use tasks. Command A (15.2%) is next.

What's a good alternative to North Mini Code (free)?

Command A (15.2%) is the closest alternative on this metric. See the full ranking above for the tradeoffs.

How many Cohere models are there?

modelgrep tracks 5 Cohere models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by North Mini Code (free). 2 of them qualify for this ranking.

More Cohere rankings

All rankings