Cheapest models for context
Ranked by input $/Mtoken, then output. Large context windows for long docs and repos.
- 1Hy3 (free)
Tencent · 262,144 context
Input
$0/M
Output
$0/M
- 2Qwen3 Coder 480B A35B (free)
Qwen · 1,048,576 context
Input
$0/M
Output
$0/M
- 3Claude Sonnet 5 (Free)
Anthropic · 1,000,000 context
Input
$0/M
Output
$0/M
- 4GLM 5.2 (Free)
Z Ai · 1,000,000 context
Input
$0/M
Output
$0/M
- 5Kimi K2.7 Code (Free)
Moonshotai · 262,144 context
Input
$0/M
Output
$0/M
- 6Nemotron 3 Ultra (free)
NVIDIA · 1,000,000 context
Input
$0/M
Output
$0/M
- 7glm-4.7
Novita · 205,000 context
Input
$0/M
Output
$0/M
- 8kimi-k2-thinking
Novita · 256,000 context
Input
$0/M
Output
$0/M
- 9minimax-m2.1
Novita · 205,000 context
Input
$0/M
Output
$0/M
- 10Step 3.7 Flash (Free)
Stepfun · 256,000 context
Input
$0/M
Output
$0/M
- 11Gemma 4 31B (free)
Google · 262,144 context
Input
$0/M
Output
$0/M
- 12GLM-4.6
Zhipuai · 202,752 context
Input
$0/M
Output
$0/M
- 13Gemma 4 26B A4B (free)
Google · 262,144 context
Input
$0/M
Output
$0/M
- 14Nemotron 3 Super (free)
NVIDIA · 1,000,000 context
Input
$0/M
Output
$0/M
- 15gpt-oss-120b (free)
OpenAI · 131,072 context
Input
$0/M
Output
$0/M
- 16glm-4.7-flash
Novita · 200,000 context
Input
$0/M
Output
$0/M
- 17GLM 4.7 Flash (Free)
Z Ai · 200,000 context
Input
$0/M
Output
$0/M
- 18Cohere Command A
Cohere · 128,000 context
Input
$0/M
Output
$0/M
- 19North Mini Code (free)
Cohere · 256,000 context
Input
$0/M
Output
$0/M
- 20GLM-4.5
Zhipuai · 131,072 context
Input
$0/M
Output
$0/M
- 21Grok 3
xAI · 128,000 context
Input
$0/M
Output
$0/M
- 22Meituan/Longcat-Flash-Lite
Meituan · 256,000 context
Input
$0/M
Output
$0/M
- 23glm-4.6v
Novita · 131,000 context
Input
$0/M
Output
$0/M
- 24gpt-oss-20b (free)
OpenAI · 131,072 context
Input
$0/M
Output
$0/M
- 25Mistral Small 3.1
Mistral Ai · 128,000 context
Input
$0/M
Output
$0/M
- 26Qwen3 30B A3B 2507
Qwen · 262,144 context
Input
$0/M
Output
$0/M
- 27Nemotron 3 Nano 30B A3B (free)
NVIDIA · 256,000 context
Input
$0/M
Output
$0/M
- 28Qwen3 Next 80B A3B Instruct (free)
Qwen · 262,144 context
Input
$0/M
Output
$0/M
- 29Mistral Medium 3 (25.05)
Mistral Ai · 128,000 context
Input
$0/M
Output
$0/M
- 30Llama 3.3 70B Instruct (free)
Meta · 131,072 context
Input
$0/M
Output
$0/M
- 31nvidia-nemotron-nano-9b-v2
NVIDIA · 131,072 context
Input
$0/M
Output
$0/M
- 32Phi-4-mini-instruct
Microsoft · 128,000 context
Input
$0/M
Output
$0/M
- 33Phi-4-Reasoning
Microsoft · 128,000 context
Input
$0/M
Output
$0/M
- 34Phi-4-multimodal-instruct
Microsoft · 128,000 context
Input
$0/M
Output
$0/M
- 35Mistral Large 24.11
Mistral Ai · 128,000 context
Input
$0/M
Output
$0/M
- 36Mistral Large 3 675B Instruct 2512
Mistral · 262,144 context
Input
$0/M
Output
$0/M
- 37Cohere Command R 08-2024
Cohere · 128,000 context
Input
$0/M
Output
$0/M
- 38Cohere Command R+ 08-2024
Cohere · 128,000 context
Input
$0/M
Output
$0/M
- 39Cohere Command R+
Cohere · 128,000 context
Input
$0/M
Output
$0/M
- 40Cohere Command R
Cohere · 128,000 context
Input
$0/M
Output
$0/M