Hermes 4 70B

Nous Research Hermes 4 hybrid reasoning model based on Llama 3.1 70B.

hermes-4-70b
STABLEGet StartedView uptime
131,072 context
Starting at $0.13/M input tokens
Starting at $0.40/M output tokens
Streaming
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

Select Provider

All Providers for Hermes 4 70B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Nebius AI
Context: 131.1kQuant: fp8
Input
$0.13
/M tokens
Cache Read
/M tokens
Output
$0.4
/M tokens
Get Started