modelgrep

Best LLMs for Tool Calling — MiniMax

Match · Updated August 2026

The best MiniMax model for tool calling is MiniMax M2.5, completing 95.3% of Tau²-Bench's multi-turn tool-use tasks. MiniMax M3 (88.9%) and MiniMax M2 (86.8%) round out the top three.

95.3%τ²-Bench
33.7Intelligence
$0.150Input /M
205KContext
  1. 1minimax logo
    minimax-m2.5
    ReasoningToolsJSON33.7 intel · $0.150/M · 205K ctx
    95.3%
    τ²-Bench
  2. 2minimax logo
    minimax-m3
    ReasoningToolsJSON+144.4 intel · $0.300/M · 101 t/s
    88.9%
    τ²-Bench
  3. 3minimax logo
    minimax-m2
    ReasoningToolsJSON28.3 intel · $0.255/M · 48 t/s
    86.8%
    τ²-Bench
  4. 4minimax logo
    minimax-m2.1
    ReasoningToolsJSON31.4 intel · $0.300/M · 63 t/s
    85.4%
    τ²-Bench
  5. 5minimax logo
    minimax-m2.7
    ReasoningToolsJSON38.1 intel · $0.270/M · 205K ctx
    84.8%
    τ²-Bench

How this is ranked

AI models ranked by Tau²-Bench — multi-turn conversations where the model has to call the right tools, in the right order, against a real API to complete a customer task. This measures whether function calling actually works under pressure, which is a different question from whether a model supports the parameter at all.

Frequently asked

Which MiniMax model is best at tool calling?

The best MiniMax model for tool calling is MiniMax M2.5, completing 95.3% of Tau²-Bench's multi-turn tool-use tasks. MiniMax M3 (88.9%) and MiniMax M2 (86.8%) round out the top three.

What's a good alternative to MiniMax M2.5?

MiniMax M3 (88.9%) is the closest alternative on this metric, followed by MiniMax M2 (86.8%). See the full ranking above for the tradeoffs.

How many MiniMax models are there?

modelgrep tracks 8 MiniMax models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by MiniMax M3. 5 of them qualify for this ranking.

More MiniMax rankings

All rankings