Nous
·Released on Aug 16, 2024Hermes 3 405B Instruct API Benchmarks, Pricing & Provider Data
Compare Hermes 3 405B Instruct with another model
Choose a model to open its comparison page.
LLM
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coheren...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
INPUT
131.1Ktokens
≈ 157.3 pages of text
OUTPUT
16.4Ktokens
8K128K1M4M
131.1K
Features
Technical Details
- Input
- Output
- Total parameters
- 405B
- Released
- Aug 2024
- Knowledge cutoff
- 2023-12-31
- Tokenizer
- Llama3
- Architecture
- text->text
- Instruct type
- chatml
- Moderated
- No
- Supported parameters
- frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p
OpenRouter endpoints
1 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
DeepInfra deepinfra/fp8 | $1/M | $1/M | 100% | — | — | undefined tokens / undefined tokens |
