weightsapi.INFERENCEConsole
Navigation
Qwen3 32B vs DeepSeek V4 Flash

Qwen3 32B vs DeepSeek V4 Flash

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
QQwen3 32B$0.15$0.20128K*Text
DDeepSeek V4 Flash$0.09$0.18128K*Text

Qwen3 32B

A balanced choice for code, multilingual chat and reasoning.

Native context
32,768 tokens
Quantization
Not verified
License
Apache 2.0
Output speed
No measurements yet
Read model details ↗

DeepSeek V4 Flash

A text model for long-context reasoning, coding and tool workflows, with a configured 1,048,576-token window.

Published context
1,048,576 tokens
Quantization
Not verified
License
MIT
Output speed
No measurements yet
Read model details ↗