weightsapi.INFERENCEConsole
Navigation
Qwen3 235B A22B vs Llama 3.1 405B

Qwen3 235B A22B vs Llama 3.1 405B

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
QQwen3 235B A22B$0.20$0.60128K*Text
LLlama 3.1 405B$1.20$1.20128K*Text

Qwen3 235B A22B

Strong general-purpose instruction following and tool use.

Native context
262,144 tokens
Quantization
Not verified
License
Apache 2.0
Output speed
No measurements yet
Read model details ↗

Llama 3.1 405B

Demanding general-purpose generation and complex instructions.

Native context
131,072 tokens
Quantization
Not verified
License
Llama 3.1 Community
Output speed
No measurements yet
Read model details ↗