Skip to content
llmprice

Hermes 2 Pro Llama 3 8B

Updated Sep 11, 2026

llama131k context131k max outputreleased May 27, 2024Open weightsTool calling
Cheapest input
$0.14
Cheapest output
$0.14
Official
output per 1M tokens
Providers
2
offering this model
Spread
1.0×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
HeliconeCheapest$0.14Cheapest$0.14131k
NovitaAIopen$0.14$0.148k

Other llama models