Navigation
Qwen3 32B vs Llama 3.1 405B
Qwen3 32B vs Llama 3.1 405B
Compare the exact deployment revisions, published rates and context limits.
Generated from the model catalogue
Values here, in pricing and in the machine-readable feed use the same model records.
| Model | Input / 1M | Output / 1M | Context | Type |
|---|---|---|---|---|
| Qwen3 32B | $0.15 | $0.20 | 128K* | Text |
| Llama 3.1 405B | $1.20 | $1.20 | 128K* | Text |
Qwen3 32B
A balanced choice for code, multilingual chat and reasoning.
- Native context
- 32,768 tokens
- Quantization
- Not verified
- License
- Apache 2.0
- Output speed
- No measurements yet
Llama 3.1 405B
Demanding general-purpose generation and complex instructions.
- Native context
- 131,072 tokens
- Quantization
- Not verified
- License
- Llama 3.1 Community
- Output speed
- No measurements yet