Atlas / Models / Nousresearch / Nous: Hermes 4 405B

Nous: Hermes 4 405B✓ Catalog verified

nousresearch/hermes-4-405b

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with <think>...</think> traces or respond directly, offering flexibility between speed and depth. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. Learn more in our docs The model is instruction-tuned with an expanded post-training corpus (~60B tokens) emphasizing reasoning traces, improving performance in math, code, STEM, and logical reasoning, while retaining broad assistant utility. It also supports structured outputs, including JSON mode, schema adherence, function calling, and tool use. Hermes 4 is trained for steerability, lower refusal rates, and alignment toward neutral, user-directed behavior.

Input price
$1.00 /1M
Output price
$3.00 /1M
Context
131K
Modalities
text
Released
Aug 26, 2025
Tool calling
Atlas signal
92/100
01

Overview

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with <think>...</think> traces or respond directly, offering flexibility between speed and depth. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. Learn more in our docs The model is instruction-tuned with an expanded post-training corpus (~60B tokens) emphasizing reasoning traces, improving performance in math, code, STEM, and logical reasoning, while retaining broad assistant utility. It also supports structured outputs, including JSON mode, schema adherence, function calling, and tool use. Hermes 4 is trained for steerability, lower refusal rates, and alignment toward neutral, user-directed behavior.

Access: available through the official Nousresearch API. Context window: 131K.

Pricing and metadata from the ModelsAtlas catalog.Last refreshed Aug 5, 2026
02

Specifications

API identifiernousresearch/hermes-4-405b
ProviderNousresearch
Model typeText LLM
Context window131K catalog
Input modalitiestext
Output modalitiestext
ReleasedAug 26, 2025
ModeratedNo
Architecture modalitytext->text
Supported parameters
frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningrepetition_penaltyresponse_formattemperaturetop_ktop_p
Source: provider documentation + catalog feedMethodology →
03

Pricing

Live pricing components for nousresearch/hermes-4-405b as published in the catalog.

$1.00
Input /1M
$3.00
Output /1M
Per request
Per image
Pricing componentRaw unit priceNormalizedUnit
Prompt tokens$0.000001$1.00 / 1MPer input token
Completion tokens$0.000003$3.00 / 1MPer output token
Request feeN/AN/APer request
Image feeN/AN/APer image unit
Web search feeN/AN/APer search request
Source: catalog pricing feedCompare all pricing →Cheapest models →
04

Cost calculator

Estimate monthly spend from your own token volumes.

Scale / Volume

Estimated Total Cost

$1.1Calculated from current list pricing in the ModelsAtlas catalog.
05

Capabilities

CapabilityStatusWhat it means
Visual UnderstandingNot advertisedImage and document analysis support
Audio ProcessingNot advertisedSpeech and voice aligned flows
Tool CallingNot advertisedSupports tools / function calling
Self-HostingNot advertisedDeploy outside managed APIs
Derived from catalog capability tags and supported parameters
06

Capability signals

Directional signals derived from model metadata and capability tags — not official benchmark submissions.

Nous: Hermes 4 405B
MMLU Signal (Reasoning)
97signal
Coding Signal (HumanEval proxy)
92signal
Math Signal (GSM8K proxy)
92signal
Science Signal (GPQA proxy)
87signal
Estimated from metadata — not an official benchmark runAll benchmarks →Methodology →
07

Quick start

Call Nous: Hermes 4 405B through an OpenAI-compatible client.

from openai import OpenAI

client = OpenAI(
    base_url="https://openrouter.ai/api/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="nousresearch/hermes-4-405b",
    messages=[{"role": "user", "content": "Explain quantum physics."}]
)

print(response.choices[0].message.content)
08

Alternatives

Closest models by context window from a different provider.

09

Sources & attribution

Pricing and metadata are maintained in the ModelsAtlas catalog.

Capability bars use metadata tags — directional estimates onlyHow we source and verify data →
10

Frequently asked questions

How much does Nous: Hermes 4 405B cost?

$1.00 per 1M input tokens and $3.00 per 1M output tokens on the official Nousresearch API.

What is the context window of Nous: Hermes 4 405B?

Nous: Hermes 4 405B supports up to 131K tokens of context.

Does Nous: Hermes 4 405B support tool / function calling?

Nous: Hermes 4 405B does not advertise first-class tool/function calling. Check vendor docs for the latest capabilities.

How do I access Nous: Hermes 4 405B?

Use the official Nousresearch API with the model id `nousresearch/hermes-4-405b`. See the Quick start section above for code examples.