weightsapi.INFERENCEConsole
Navigation
Llama 3.1 8B vs Dolphin Mistral 24B Venice Edition

Llama 3.1 8B vs Dolphin Mistral 24B Venice Edition

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
LLlama 3.1 8B$0.03$0.05128K*Text
DDolphin Mistral 24B Venice Edition$0.20$0.90128K*Text + image

Llama 3.1 8B

Fast classification, extraction and lightweight chat.

Native context
131,072 tokens
Quantization
Not verified
License
Llama 3.1 Community
Output speed
No measurements yet
Read model details ↗

Dolphin Mistral 24B Venice Edition

Conversations and characters with application-controlled instructions, from Dolphin and Venice.

Published context
131,072 tokens
Quantization
Not verified
License
Apache 2.0
Output speed
No measurements yet
Read model details ↗