Currently listed through:
Model details
Hermes 3 70B Instruct
Hermes 3 70B Instruct is a 70-billion-parameter fine-tune of the Llama 3.1 70B foundation model, released by Nous Research as a generalist chat model with an emphasis on giving users strong steering control. The model is designed to feel less constrained than typical assistant-tuned LLMs, prioritizing alignment to the end user and exposing capabilities such as advanced agentic behavior, richer roleplay, deeper multi-turn conversation, and long-context coherence. Sources describe it as a "competitive, if not superior" fine-tune focused on powerful, reliable function calling, structured output, and improved code generation, making it a flexible base for both conversational interfaces and tool-driven workflows.
Building on the Hermes 2 lineage, Hermes 3 broadens the family with more capable agentic function calling, structured outputs, and better generalist assistant behavior across reasoning and coding tasks. Independent benchmark coverage places the model in the upper tier of evaluated open-weight systems, with standout performance on instruction following (IFEval 76.6%) and broad reasoning (BBH 53.8%), alongside competitive results on knowledge-heavy suites. Because the weights are openly published, practitioners can self-host the model or run it through third-party endpoints, which makes it a practical option for customized assistants, agent pipelines, and long-context applications where control over model behavior is a priority.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- nousresearch/hermes-3-llama-3.1-70b
- Release date
- Aug 18, 2024
- Last updated
- Aug 18, 2024
- Knowledge cutoff
- 2023-12-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.70
- Output token cost
- $0.70
Limits
- Output tokens
- 16,384 tokens
- Context window
- 131,072 tokens