Provider logo

Hermes 4 Large

nousresearch/hermes-4-405b
Provider logo

Hermes 4 Large

nousresearch/hermes-4-405b

Advanced reasoning model built on Llama-3.1-405B with hybrid thinking modes. Features internal deliberation capabilities, excels at math, code, STEM, and logical reasoning while supporting structured outputs with improved steerability and neutral alignment.

Added Aug 26, 2025

Model weights

Context Window

128.0K

Max Output

8.2K

Input Price (Auto)

$0.30/1M

Output Price (Auto)

$1.20/1M

Cache Read (Auto)

$0.15/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

8.6

Better than 26% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

53.6%

Better than 29% of models compared

HLE

Humanity's Last Exam

4.2%

Better than 18% of models compared

IFBench

Instruction-following benchmark

34.8%

Better than 26% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

26.6%

Better than 32% of models compared

AA-LCR

Long context reasoning evaluation

21.0%

Better than 29% of models compared

Coding

SciCode

Python programming for scientific computing

34.6%

Better than 51% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

9.8%

Better than 44% of models compared

LiveCodeBench

Contamination-free coding benchmark

54.6%

Better than 62% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

15.3%

Better than 19% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

72.9%

Better than 42% of models compared

Last updated Aug 16, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Hermes 4 Large with similar models from the same provider or model family.