Provider logo

Hermes 4 Medium

nousresearch/hermes-4-70b
Provider logo

Hermes 4 Medium

nousresearch/hermes-4-70b

Efficient reasoning model based on Llama-3.1-70B. Offers hybrid thinking capabilities with strong performance in math, code, and logical reasoning tasks. Supports structured outputs and JSON mode with enhanced steerability.

Added Jul 3, 2025

Model weights

Context Window

128.0K

Max Output

8.2K

Input Price (Auto)

$0.20/1M

Output Price (Auto)

$0.40/1M

Cache Read (Auto)

$0.10/1M

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

6.7

Better than 19% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

49.1%

Better than 24% of models compared

HLE

Humanity's Last Exam

3.6%

Better than 7% of models compared

IFBench

Instruction-following benchmark

29.0%

Better than 13% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

21.6%

Better than 22% of models compared

AA-LCR

Long context reasoning evaluation

3.0%

Better than 14% of models compared

Coding

SciCode

Python programming for scientific computing

27.7%

Better than 35% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

0.0%

Better than 5% of models compared

LiveCodeBench

Contamination-free coding benchmark

26.9%

Better than 29% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

11.3%

Better than 15% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

66.4%

Better than 29% of models compared

Last updated Aug 16, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Hermes 4 Medium with similar models from the same provider or model family.