Navigation
Llama 3.1 405B vs Mistral Small 4
Llama 3.1 405B vs Mistral Small 4
Compare the exact deployment revisions, published rates and context limits.
Generated from the model catalogue
Values here, in pricing and in the machine-readable feed use the same model records.
| Model | Input / 1M | Output / 1M | Context | Type |
|---|---|---|---|---|
| Llama 3.1 405B | $1.20 | $1.20 | 128K* | Text |
| Mistral Small 4 | $0.15 | $0.60 | 128K* | Text + image |
Llama 3.1 405B
Demanding general-purpose generation and complex instructions.
- Native context
- 131,072 tokens
- Quantization
- Not verified
- License
- Llama 3.1 Community
- Output speed
- No measurements yet
Mistral Small 4
An official Mistral model combining text and image input, tool calling and configurable reasoning, with a published 262,144-token window.
- Published context
- 262,144 tokens
- Quantization
- Not verified
- License
- Apache 2.0
- Output speed
- No measurements yet