Cheapest models for reason
Ranked by input $/Mtoken, then output. Hard reasoning and planning workloads.
- 1Claude Sonnet 5 (Free)
Anthropic · 1,000,000 context
Input
$0/M
Output
$0/M
- 2GLM 5.2 (Free)
Z Ai · 1,000,000 context
Input
$0/M
Output
$0/M
- 3Kimi K2.7 Code (Free)
Moonshotai · 262,144 context
Input
$0/M
Output
$0/M
- 4glm-4.7
Novita · 205,000 context
Input
$0/M
Output
$0/M
- 5kimi-k2-thinking
Novita · 256,000 context
Input
$0/M
Output
$0/M
- 6minimax-m2.1
Novita · 205,000 context
Input
$0/M
Output
$0/M
- 7Step 3.7 Flash (Free)
Stepfun · 256,000 context
Input
$0/M
Output
$0/M
- 8GLM-4.6
Zhipuai · 202,752 context
Input
$0/M
Output
$0/M
- 9gpt-oss-120b (free)
OpenAI · 131,072 context
Input
$0/M
Output
$0/M
- 10glm-4.7-flash
Novita · 200,000 context
Input
$0/M
Output
$0/M
- 11GLM 4.7 Flash (Free)
Z Ai · 200,000 context
Input
$0/M
Output
$0/M
- 12Cohere Command A
Cohere · 128,000 context
Input
$0/M
Output
$0/M
- 13GLM-4.5
Zhipuai · 131,072 context
Input
$0/M
Output
$0/M
- 14Grok 3
xAI · 128,000 context
Input
$0/M
Output
$0/M
- 15glm-4.6v
Novita · 131,000 context
Input
$0/M
Output
$0/M
- 16Mistral Small 3.1
Mistral Ai · 128,000 context
Input
$0/M
Output
$0/M
- 17Mistral Medium 3 (25.05)
Mistral Ai · 128,000 context
Input
$0/M
Output
$0/M
- 18nvidia-nemotron-nano-9b-v2
NVIDIA · 131,072 context
Input
$0/M
Output
$0/M
- 19Phi-4-mini-instruct
Microsoft · 128,000 context
Input
$0/M
Output
$0/M
- 20Phi-4-Reasoning
Microsoft · 128,000 context
Input
$0/M
Output
$0/M
- 21Phi-4-multimodal-instruct
Microsoft · 128,000 context
Input
$0/M
Output
$0/M
- 22Mistral Large 24.11
Mistral Ai · 128,000 context
Input
$0/M
Output
$0/M
- 23Cohere Command R
Cohere · 128,000 context
Input
$0/M
Output
$0/M
- 24LFM2.5-1.2B-Thinking (free)
Liquid · 32,768 context
Input
$0/M
Output
$0/M
- 25AI21 Jamba 1.5 Large
Ai21 Labs · 256,000 context
Input
$0/M
Output
$0/M
- 26AI21 Jamba 1.5 Mini
Ai21 Labs · 256,000 context
Input
$0/M
Output
$0/M
- 27Baidu: CoBuddy (free)
Baidu · 131,072 context
Input
$0/M
Output
$0/M
- 28qwen3-235b-2507-cs
Cerebras · 1 context
Input
$0/M
Output
$0/M
- 29qwen3-32b-cs
Cerebras · 1 context
Input
$0/M
Output
$0/M
- 30JAIS 30b Chat
Core42 · 8,192 context
Input
$0/M
Output
$0/M
- 31Deepseek/Deepseek-Math-V2
DeepSeek · 160,000 context
Input
$0/M
Output
$0/M
- 32Deepseek/DeepSeek-V3.2
DeepSeek · 128,000 context
Input
$0/M
Output
$0/M
- 33DeepSeek/DeepSeek-V3.1-Terminus-Thinking
DeepSeek · 128,000 context
Input
$0/M
Output
$0/M
- 34DeepSeek/DeepSeek-V3.2-Exp-Thinking
DeepSeek · 128,000 context
Input
$0/M
Output
$0/M
- 35Gemma 4 31B IT FP8
Google · 262,144 context
Input
$0/M
Output
$0/M
- 36Kilo Auto Free
Kilo Auto · 204,800 context
Input
$0/M
Output
$0/M
- 37LongCat Flash Thinking 2601
Meituan · 32,768 context
Input
$0/M
Output
$0/M
- 38Llama-3.2-11B-Vision-Instruct
Meta · 128,000 context
Input
$0/M
Output
$0/M
- 39Llama-3.2-90B-Vision-Instruct
Meta · 128,000 context
Input
$0/M
Output
$0/M
- 40Llama-3.3-70B-Instruct
Meta · 128,000 context
Input
$0/M
Output
$0/M