modelgrep

Cheapest LLMs

Match · Updated July 2026

The cheapest LLM is Ling-2.6-flash at $0.010 per million input tokens. Granite 4.0 Micro ($0.017) and Mistral Nemo ($0.019) round out the top three.

$0.010Input /M
14.1Intelligence
262KContext

AI models ranked by input token price. The cheapest LLM APIs and most affordable AI models, from budget open-weight models to discounted frontier models.

  1. 1I
    ling-2.6-flash
    ToolsJSON14.1 intel · 262K ctx
    $0.010
    Input /M
  2. 2ibm-granite logo
    granite-4.0-h-micro
    JSON131K ctx
    $0.017
    Input /M
  3. 3mistralai logo
    mistral-nemo
    ToolsJSON131K ctx
    $0.019
    Input /M
  4. 4N
    nex-n2-mini
    ReasoningToolsJSON+1262K ctx
    $0.025
    Input /M
  5. 5openai logo
    gpt-5-nano:batch
    ReasoningToolsJSON+119.9 intel · 400K ctx
    $0.025
    Input /M
  6. 6meta-llama logo
    llama-3.2-1b-instruct
    60K ctx
    $0.027
    Input /M
  7. 7qwen logo
    qwen3.7-flash
    ReasoningToolsJSON+11M ctx
    $0.030
    Input /M
  8. 8openai logo
    gpt-oss-20b
    ReasoningToolsJSON14.9 intel · 131K ctx
    $0.030
    Input /M
  9. 9amazon logo
    nova-micro-v1
    Tools128K ctx
    $0.035
    Input /M
  10. 10openai logo
    gpt-oss-120b
    ReasoningToolsJSON23.8 intel · 131K ctx
    $0.037
    Input /M
  11. 11cohere logo
    command-r7b-12-2024
    JSON128K ctx
    $0.037
    Input /M
  12. 12S
    l3-lunaris-8b
    JSON8K ctx
    $0.040
    Input /M
  13. 13qwen logo
    qwen3-30b-a3b-instruct-2507
    ToolsJSON262K ctx
    $0.048
    Input /M
  14. 14ibm-granite logo
    granite-4.1-8b
    ToolsJSON6.7 intel · 131K ctx
    $0.050
    Input /M
  15. 15nvidia logo
    nemotron-3-nano-30b-a3b
    ReasoningToolsJSON7.4 intel · 262K ctx
    $0.050
    Input /M
  16. 16openai logo
    gpt-5-nano
    ReasoningToolsJSON+119.9 intel · 400K ctx
    $0.050
    Input /M
  17. 17google logo
    gemini-2.5-flash-lite:batch
    ReasoningToolsJSON+26.9 intel · 1.0M ctx
    $0.050
    Input /M
  18. 18google logo
    gemma-3-4b-it
    JSONVision131K ctx
    $0.050
    Input /M
  19. 19google logo
    gemma-3-12b-it
    ToolsJSONVision131K ctx
    $0.050
    Input /M
  20. 20mistralai logo
    mistral-small-24b-instruct-2501
    JSON33K ctx
    $0.050
    Input /M
  21. 21meta-llama logo
    llama-3.2-3b-instruct
    JSON131K ctx
    $0.050
    Input /M
  22. 22meta-llama logo
    llama-3.1-8b-instruct
    ToolsJSON131K ctx
    $0.050
    Input /M
  23. 23P
    laguna-xs-2.1
    ReasoningTools262K ctx
    $0.060
    Input /M
  24. 24z-ai logo
    glm-4.7-flash
    ReasoningToolsJSON22.9 intel · 203K ctx
    $0.060
    Input /M
  25. 25google logo
    gemma-3n-e4b-it
    JSON33K ctx
    $0.060
    Input /M

Frequently asked

What is the cheapest LLM?

The cheapest LLM is Ling-2.6-flash at $0.010 per million input tokens. Granite 4.0 Micro ($0.017) and Mistral Nemo ($0.019) round out the top three.

What is the cheapest AI API?

The cheapest AI model is Ling-2.6-flash at $0.010 per million input tokens. Granite 4.0 Micro ($0.017) and Mistral Nemo ($0.019) round out the top three.

What's a good alternative to Ling-2.6-flash?

Granite 4.0 Micro ($0.017) is the closest alternative on this metric, followed by Mistral Nemo ($0.019). See the full ranking above for the tradeoffs.

By maker

All rankings