Navigation
Qwen3 32B vs Llama 3.3 70B
Qwen3 32B vs Llama 3.3 70B
Compare the exact deployment revisions, published rates and context limits.
Generated from the model catalogue
Values here, in pricing and in the machine-readable feed use the same model records.
| Model | Input / 1M | Output / 1M | Context | Type |
|---|---|---|---|---|
| Qwen3 32B | $0.15 | $0.20 | 128K* | Text |
| Llama 3.3 70B | $0.35 | $0.40 | 128K* | Text |
Qwen3 32B
A balanced choice for code, multilingual chat and reasoning.
- Native context
- 32,768 tokens
- Quantization
- Not verified
- License
- Apache 2.0
- Output speed
- No measurements yet
Llama 3.3 70B
General-purpose assistants, summarization and tool use.
- Native context
- 131,072 tokens
- Quantization
- Not verified
- License
- Llama 3.3 Community
- Output speed
- No measurements yet