Cheapest capable models

Live list prices across providers. Ranking here is arithmetic, not opinion.

Right now

Mistral Nemo from Mistral currently ranks first for cost per million tokens, ahead of Ling-3.0-flash from —. The ranking comes from OpenRouter and was last published 2026-09-02.

Leader
Mistral Nemo
Made by
Mistral
Blended $/1M
$0.022
Source
OpenRouter

Full ranking Published list prices, blended 3:1 input to output.

Rank# Model Made by Per 1M tokens Context
1 Mistral Nemo Mistral $0.022 131K
2 Ling-3.0-flash $0.032 262K
3 Granite 4.0 Micro IBM $0.041 131K
4 Nex-N2-Mini Nex AGI $0.044 262K
5 Solar Pro 4 Upstage $0.052 524K
6 Qwen3.7 Flash Qwen $0.055 1M
7 gpt-oss-20b OpenAI $0.055 131K
8 Llama 3.1 8B Instruct Meta $0.057 131K
9 Nova Micro 1.0 Amazon $0.061 128K
10 Granite 4.1 8B IBM $0.062 131K
11 Gemma 3 4B Google $0.062 131K
12 Command R7B (12-2024) Cohere $0.066 128K
13 Mercury 2.5 Preview Inception $0.068 260K
14 GPT-5 Nano (batch) OpenAI $0.069 400K
15 gpt-oss-120b OpenAI $0.070 131K
16 Laguna XS 2.1 Poolside $0.075 262K
17 Gemma 3 12B Google $0.075 131K
18 DeepSeek V4 Flash Latest $0.077 1.311M
19 Qwen3 30B A3B Instruct 2507 Qwen $0.084 262K
20 Nemotron 3 Nano 30B A3B NVIDIA $0.087 262K
21 gpt-oss-20b (batch) OpenAI $0.087 131K
22 Gemini 2.5 Flash Lite (batch) Google $0.087 1.049M
23 GPT-4.1 Nano (batch) OpenAI $0.087 1.048M
24 DeepSeek V4 Flash 0731 DeepSeek $0.094 1.311M
25 DeepSeek V4 Flash 0423 DeepSeek $0.099 1.049M

Published by OpenRouter on 2026-09-02 · copied here 2026-09-02 · no adjustment applied

Other leaderboards