Nous Research Hermes 2 Pro based on Llama 3 8B.
OffRail routes requests to the best providers that are able to handle your prompt size and parameters.
Nous Research Hermes 2 Pro based on Llama 3 8B. You can access it through OffRail's OpenAI-compatible API with automatic provider routing, fallback, and cost analytics.
Pricing for Hermes 2 Pro Llama 3 8B on OffRail starts at $0.14 per million input tokens and $0.14 per million output tokens, depending on the provider. The pricing table above always reflects the current per-provider rates.
Hermes 2 Pro Llama 3 8B supports a context window of up to 8,192 tokens on its largest provider deployment.
Hermes 2 Pro Llama 3 8B is served by NovitaAI through OffRail. Requests are automatically routed to the best available provider, with fallback when a provider has issues.
No. Hermes 2 Pro Llama 3 8B does not currently support tool calling or structured JSON outputs through OffRail.
Hermes 2 Pro Llama 3 8B was released on May 27, 2024.