Nvidia Nemotron 3.5 Lightning TEE
NVIDIA's open-weight 30B mixture-of-experts model with 3B active parameters for fast agentic workflows, coding, and tool use. Running inside a TEE (Trusted Execution Environment), with provider attestation support.
- Reasoning
- Tool Calling
- Structured Output
Added Sep 11, 2026
Model weightsPricing
Auto routing · per 1M tokens- Input
- $0.080
- Output
- $0.20
- Cache read
- $0.040
Specifications
- Context window
- 262.1K
- Max output
- 65.5K
- Parameters
- 30B / 3B
- Total / active
- Avg output (7d)
- 507 tokens
- Longer than 44% of models
Benchmarks
Benchmarks
Sourced from Artificial Analysis.
Intelligence Index
12.9
Coding Index
26.8
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
0.8%
Better than 12% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
505 Elo
Better than 24% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
624 Elo
Better than 32% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
3.0%
Better than 17% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
60.3%
Better than 51% of models compared
MLCR-AA
Medical long-context reasoning
1.1%
Better than 7% of models compared
Reasoning
HLE
Humanity's Last Exam
10.6%
Better than 52% of models compared
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
0.5%
Better than 35% of models compared
SciCode
Python programming for scientific computing
32.1%
Better than 11% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
74.3%
Better than 55% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
60.3%
Better than 51% of models compared
Last updated Oct 1, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…