weightsapi.INFERENCEConsole
Navigation
Mixtral 8×22B vs Llama 3.1 405B

Mixtral 8×22B vs Llama 3.1 405B

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
MMixtral 8×22B$0.60$0.6064K*Text
LLlama 3.1 405B$1.20$1.20128K*Text

Mixtral 8×22B

Multilingual generation and function-calling workflows.

Native context
65,536 tokens
Quantization
Not verified
License
Apache 2.0
Output speed
No measurements yet
Read model details ↗

Llama 3.1 405B

Demanding general-purpose generation and complex instructions.

Native context
131,072 tokens
Quantization
Not verified
License
Llama 3.1 Community
Output speed
No measurements yet
Read model details ↗