Hermes 3 70B Instruct pricing & benchmarks
Released Aug 18, 2024 · served by 1 provider
What Hermes 3 70B Instruct is
Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
- Model ID:
nousresearch/hermes-3-llama-3.1-70b - Modality: text->text
- Knowledge cutoff: 2023-12-31
- Tool calling: not supported · Structured output: supported
- Weights:
NousResearch/Hermes-3-Llama-3.1-70Bon Hugging Face
Providers and prices for Hermes 3 70B Instruct
The same model costs different amounts depending on who serves it. Prices are per 1M tokens, cheapest first. Uptime and throughput are OpenRouter's rolling measurements, not vendor claims.
| Provider | Input | Output | Context | Quant | Uptime 24h | Throughput |
|---|---|---|---|---|---|---|
| DeepInfracheapest | $0.700 | $0.700 | 131K | fp8 | 100.0% | — |
FAQ
How much does the Hermes 3 70B Instruct API cost?
$0.700 per 1M input tokens and $0.700 per 1M output tokens. A typical workload of 1M input + 300K output tokens costs about $0.91.
Is Hermes 3 70B Instruct free?
No. It is a paid model starting at $0.700 per 1M input tokens, though some providers offer trial credits.
Which provider is cheapest for Hermes 3 70B Instruct?
DeepInfra at $0.700 input / $0.700 output per 1M tokens (fp8 quantization).
Can I self-host Hermes 3 70B Instruct?
Yes — weights are published as NousResearch/Hermes-3-Llama-3.1-70B on Hugging Face, so you can run it on your own hardware.