weightsapi.INFERENCEConsole
Navigation
Llama 3.1 8B vs Hermes 4 70B

Llama 3.1 8B vs Hermes 4 70B

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
LLlama 3.1 8B$0.03$0.05128K*Text
HHermes 4 70B$0.13$0.40128K*Text

Llama 3.1 8B

Fast classification, extraction and lightweight chat.

Native context
131,072 tokens
Quantization
Not verified
License
Llama 3.1 Community
Output speed
No measurements yet
Read model details ↗

Hermes 4 70B

Creative conversation, roleplay and reasoning with configurable instructions, published by Nous Research.

Published context
131,072 tokens
Quantization
Not verified
License
Meta Llama Community
Output speed
No measurements yet
Read model details ↗