Meta · open weights
Llama 3.3 70B Instruct API pricing by provider
Llama 3.3 70B Instruct is available from 10 providers. The cheapest is DeepInfra at $0.10 input / $0.32 output per 1M tokens (FP8). The priciest host, Together, charges 6.7x the cheapest.
Prices from the OpenRouter API, retrieved September 15, 2026. Weights: meta-llama/Llama-3.3-70B-Instruct.
| Provider | Precision | Input / 1M | Output / 1M | Document Q&A / month | Context | Max output |
|---|---|---|---|---|---|---|
| DeepInfra | fp8 | $0.10 | $0.32 | $49.60 | 131K | 16K |
| Novita | bf16 | $0.135 | $0.40 | $66.00 | 12K | 11K |
| AkashML | fp8 | $0.20 | $0.52 | $95.60 | 131K | 128K |
| Parasail | fp8 | $0.22 | $0.50 | $103.00 | 131K | 16K |
| SambaNova | — | $0.45 | $0.90 | $207.00 | 131K | 3K |
| Groq | — | $0.59 | $0.79 | $259.70 | 131K | 33K |
| CoreWeave | fp16 | $0.71 | $0.71 | $305.30 | 128K | 115K |
| — | $0.72 | $0.72 | $309.60 | 128K | 8K | |
| Cloudflare | fp8 | $0.293 | $2.253 | $184.79 | 24K | 22K |
| Together | — | $1.04 | $1.04 | $447.20 | 131K | 2K |
Sorted by blended price (3 input : 1 output). Document Q&A assumes 50,000 questions a month with 8,000 input and 600 output tokens each. Precision is reported by the provider; “—” means not disclosed.
Frequently asked questions
- What is the cheapest Llama 3.3 70B Instruct API provider?
- DeepInfra is the cheapest healthy provider we track at $0.10 per 1M input tokens and $0.32 per 1M output tokens.
- How much does Llama 3.3 70B Instruct cost?
- Across 10 providers, input prices range from $0.10 to $1.04 per 1M tokens and output prices from $0.32 to $2.253.
- What is Llama 3.3 70B Instruct’s context window?
- Llama 3.3 70B Instruct supports up to 131K tokens, but some providers serve a smaller context window — check the table before choosing.
Other open-weight models
- DeepSeek V4.1 Flashfrom $0.15 / $0.60
- DeepSeek V4 Pro 0813from $0.96 / $2.88
- Kimi K3from $2.10 / $10.95
- GLM 5.3from $0.91 / $2.86
- GLM 5.3 Flashfrom $0.075 / $0.25
- Qwen3.8 27Bfrom $0.15 / $2.00
- MiniMax M3from $0.23 / $0.96
- gpt-oss-120bfrom $0.03 / $0.17
- gpt-oss-20bfrom $0.02 / $0.10
- Llama 4 Maverickfrom $0.1875 / $0.6525
- Gemma 4 31Bfrom $0.09 / $0.34
- Mistral Small 4from $0.15 / $0.60
- Nemotron 3 Superfrom $0.085 / $0.40