Nvidia Nemotron 70b
Nvidia's latest Llama fine-tune optimized for instruction following. Early results hints that it might outperform models such as GPT-4o and Claude 3.5 Sonnet.
Added Apr 15, 2025
Pricing
Auto routing · per 1M tokens- Input
- $0.36
- Output
- $0.41
Specifications
- Context window
- 16.4K
- Max output
- 8.2K
- Parameters
- 70B
- Avg output (7d)
- 328 tokens
- Longer than 29% of models
Benchmarks
Benchmarks
Sourced from Artificial Analysis.
Intelligence Index
6.9
Agentic work
T²-Bench Telecom (legacy)
Legacy fallback · Conversational AI agents in dual-control scenarios
23.1%
Better than 25% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
8.3%
Better than 14% of models compared
Reasoning
HLE
Humanity's Last Exam
4.2%
Better than 16% of models compared
IFBench
Instruction-following benchmark
30.7%
Better than 15% of models compared
Coding
Terminal-Bench Hard (legacy)
Legacy fallback · Agentic coding and terminal use
4.5%
Better than 29% of models compared
LiveCodeBench
Contamination-free coding benchmark
16.9%
Better than 18% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
11.0%
Better than 14% of models compared
AIME
American Invitational Mathematics Examination
24.7%
Better than 52% of models compared
Math-500
Diverse mathematical problem solving benchmark
73.3%
Better than 34% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
69.0%
Better than 33% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
46.5%
Better than 20% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
8.3%
Better than 14% of models compared
Last updated Oct 3, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…