Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Nous: Hermes 4 405B

Hermes 4 405B is a large-scale reasoning model built upon the Meta-Llama-3.1-405B architecture, representing a significant effort to provide high-performance language capabilities. The model is designed with a unique hybrid reasoning mode, allowing it to either deliberate internally using specific reasoning traces or provide direct responses based on user preference. This flexibility makes it well-suited for complex tasks that require a balance between rapid interaction and deep, analytical processing. Beyond its core reasoning functions, the model is engineered to support structured outputs, including JSON mode, schema adherence, and robust tool use, making it a versatile choice for developers and enterprises.

The model underwent extensive instruction tuning, incorporating an expanded post-training corpus of approximately 60 billion tokens that specifically emphasizes reasoning traces. This training approach enhances its performance across math, coding, and STEM-related domains while maintaining a neutral, user-directed tone. By focusing on steerability and reducing refusal rates, the model is optimized for reliable assistant utility in professional workflows. As an open-weight model, it serves as a powerful alternative for those seeking frontier-level reasoning capabilities without the constraints of closed-source systems, positioning it as a strong candidate for both research and commercial applications.

Kilo Gatewaynousresearch/hermes-4-405bnousresearch

Quick Info

Powered by
Provider
Kilo Gateway
Model key
nousresearch/hermes-4-405b
Release date
Aug 26, 2025
Last updated
Aug 26, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$3.00

Limits

Output tokens
117,964 tokens
Context window
131,072 tokens

Latest news about Nous: Hermes 4 405B

Kilo Gateway

CoverageBenchmark

Nous Research's Hermes 4 405B is a large-scale reasoning model built on Meta-Llama-3.1-405B, released on August 26, 2025 with a 131K context window and an August 2024 knowledge cutoff. It introduces a hybrid reasoning mode in which the model can either deliberate internally via ... traces or respond directly, with a co The model is trained for lower refusal rates, steerability, and neutral, user-directed alignment, and supports structured outputs including JSON mode, schema adherence, function calling, and tool use. Reported Artificial Analysis benchmark scores include GPQA Diamond 72.7%, HLE 10.9%, IFBench 32.7%, τ²-Bench Telecom 22

Videos about Nous: Hermes 4 405B

More models around Nous: Hermes 4 405B