Efficient reasoning model based on Llama-3.1-70B. Offers hybrid thinking capabilities with strong performance in math, code, and logical reasoning tasks. Supports structured outputs and JSON mode with enhanced steerability.
Added Jul 3, 2025
Model weightsContext Window
128.0K
Max Output
8.2K
Input Price (Auto)
$0.20/1M
Output Price (Auto)
$0.40/1M
Cache Read (Auto)
$0.10/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
6.7
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
49.1%
Better than 24% of models compared
HLE
Humanity's Last Exam
3.6%
Better than 7% of models compared
IFBench
Instruction-following benchmark
29.0%
Better than 13% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
21.6%
Better than 22% of models compared
AA-LCR
Long context reasoning evaluation
3.0%
Better than 14% of models compared
Coding
SciCode
Python programming for scientific computing
27.7%
Better than 35% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
0.0%
Better than 5% of models compared
LiveCodeBench
Contamination-free coding benchmark
26.9%
Better than 29% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
11.3%
Better than 15% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
66.4%
Better than 29% of models compared
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Hermes 4 Medium with similar models from the same provider or model family.
Hermes 3 70B
nousresearch/hermes-3-llama-3.1-70bHermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, better roleplaying, reasoning, multi-turn conversation, and long context coherence. This 70B model is a competitive finetune of Llama-3.1-70B focused on aligning LLMs to the user with powerful steering capabilities.
Hermes 4 (Thinking)
NousResearch/Hermes-4-70B:thinkingHermes 4 70B with thinking enabled. Emits explicit reasoning content before final answer when streamed.
Hermes 4 Large
nousresearch/hermes-4-405bAdvanced reasoning model built on Llama-3.1-405B with hybrid thinking modes. Features internal deliberation capabilities, excels at math, code, STEM, and logical reasoning while supporting structured outputs with improved steerability and neutral alignment.
Hermes 4 Large (Thinking)
nousresearch/hermes-4-405b:thinkingHermes 4 Large with thinking enabled. Streams visible reasoning before the final answer and supports structured outputs.
Hermes Medium
hermes-mediumCurrently points to MiniMax M2.7. Middle tier for Hermes-style agent work: medium intelligence and cost for capable everyday tool use.
Linkup Research Medium
linkup-research-mediumLinkup Research with medium reasoning depth. Runs an async web research agent and returns a sourced answer for factual and multi-source questions. Responses can take several minutes.