weightsapi.INFERENCEConsole
Navigation
Llama 3.1 405B vs Hermes 4 70B

Llama 3.1 405B vs Hermes 4 70B

Compare the exact deployment revisions, published rates and context limits.

Generated from the model catalogue

Values here, in pricing and in the machine-readable feed use the same model records.

ModelInput / 1MOutput / 1MContextType
LLlama 3.1 405B$1.20$1.20128K*Text
HHermes 4 70B$0.13$0.40128K*Text

Llama 3.1 405B

Demanding general-purpose generation and complex instructions.

Native context
131,072 tokens
Quantization
Not verified
License
Llama 3.1 Community
Output speed
No measurements yet
Read model details ↗

Hermes 4 70B

Creative conversation, roleplay and reasoning with configurable instructions, published by Nous Research.

Published context
131,072 tokens
Quantization
Not verified
License
Meta Llama Community
Output speed
No measurements yet
Read model details ↗