Model comparison
Llama-3.1-8B-CS vs llama-3.3-70b-cs
Compare Llama-3.1-8B-CS and llama-3.3-70b-cs: input/output $/Mtoken, context window, modalities, and license. OpenRouter-synced pricing.
Cerebras
Llama-3.1-8B-CS
Open Llama instruction model for multilingual chat, reasoning, and coding
Cerebras
llama-3.3-70b-cs
Legacy model retained for compatibility with older integrations
| Metric | Llama-3.1-8B-CS | llama-3.3-70b-cs |
|---|---|---|
| Provider | Cerebras | Cerebras |
| Context window | 128,000 | 1 |
| Input $/Mtok | $0.100 | $0 |
| Output $/Mtok | $0.100 | $0 |
| Modalities | text, vision | text, vision |
| License | closed | closed |
Quick take
On input price, llama-3.3-70b-cs is cheaper at $0/Mtok. For context window, Llama-3.1-8B-CS leads with 128,000 tokens.
Pick based on your workload: high-volume cheap inference vs long-document / agent loops. Enterprise DNA can wire either model into Omni with evals, secrets, and job orchestration.