weightsapi.INFERENCEConsole
Navigation
Llama 3.1 8B vs Mixtral 8×22B

Llama 3.1 8B vs Mixtral 8×22B

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
LLlama 3.1 8B$0.03$0.05128K*Text
MMixtral 8×22B$0.60$0.6064K*Text

Llama 3.1 8B

Fast classification, extraction and lightweight chat.

Native context
131,072 tokens
Quantization
Not verified
License
Llama 3.1 Community
Output speed
No measurements yet
Read model details ↗

Mixtral 8×22B

Multilingual generation and function-calling workflows.

Native context
65,536 tokens
Quantization
Not verified
License
Apache 2.0
Output speed
No measurements yet
Read model details ↗