Navigation
Llama 3.1 8B vs Qwen3 235B A22B
Llama 3.1 8B vs Qwen3 235B A22B
Compare the exact deployment revisions, published rates and context limits.
Generated from the model catalogue
Values here, in pricing and in the machine-readable feed use the same model records.
| Model | Input / 1M | Output / 1M | Context | Type |
|---|---|---|---|---|
| Llama 3.1 8B | $0.03 | $0.05 | 128K* | Text |
| Qwen3 235B A22B | $0.20 | $0.60 | 128K* | Text |
Llama 3.1 8B
Fast classification, extraction and lightweight chat.
- Native context
- 131,072 tokens
- Quantization
- Not verified
- License
- Llama 3.1 Community
- Output speed
- No measurements yet
Qwen3 235B A22B
Strong general-purpose instruction following and tool use.
- Native context
- 262,144 tokens
- Quantization
- Not verified
- License
- Apache 2.0
- Output speed
- No measurements yet