Model comparison
GPT-OSS-120B-CS vs llama-3.3-70b-cs
Compare GPT-OSS-120B-CS and llama-3.3-70b-cs: input/output $/Mtoken, context window, modalities, and license. OpenRouter-synced pricing.
Cerebras
GPT-OSS-120B-CS
Open-weight GPT model for self-hosted reasoning and instruction-following workloads
Cerebras
llama-3.3-70b-cs
Legacy model retained for compatibility with older integrations
| Metric | GPT-OSS-120B-CS | llama-3.3-70b-cs |
|---|---|---|
| Provider | Cerebras | Cerebras |
| Context window | 128,000 | 1 |
| Input $/Mtok | $0.350 | $0 |
| Output $/Mtok | $0.750 | $0 |
| Modalities | text, vision | text, vision |
| License | closed | closed |
Quick take
On input price, llama-3.3-70b-cs is cheaper at $0/Mtok. For context window, GPT-OSS-120B-CS leads with 128,000 tokens.
Pick based on your workload: high-volume cheap inference vs long-document / agent loops. Enterprise DNA can wire either model into Omni with evals, secrets, and job orchestration.