A fast, cost-effective reasoning model with 1M token context. Supports extended thinking with adjustable depth levels and built-in web grounding and code interpreter tools.
Context Window
1.0M
Max Output
65.5K
Input Price (Auto)
$0.51/1M
Output Price (Auto)
$4.25/1M
Cache Read (Auto)
$0.26/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
11.8
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
60.3%
Better than 37% of models compared
HLE
Humanity's Last Exam
2.9%
Better than 2% of models compared
IFBench
Instruction-following benchmark
40.5%
Better than 40% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
62.0%
Better than 58% of models compared
AA-LCR
Long context reasoning evaluation
18.3%
Better than 26% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
SciCode
Python programming for scientific computing
24.0%
Better than 27% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
6.8%
Better than 37% of models compared
LiveCodeBench
Contamination-free coding benchmark
34.6%
Better than 42% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
33.7%
Better than 34% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
74.3%
Better than 46% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
14.0%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
85.9%
Last updated Aug 16, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Amazon Nova 2 Lite with similar models from the same provider or model family.
Amazon Nova Lite 1.0
amazon/nova-lite-v1Amazon's new lower cost model. Can handle up to 300k input tokens, with faster output but less thorough understanding than Amazon's Nova Pro.
Amazon Nova Pro 1.0
amazon/nova-pro-v1Amazon's new flagship model. Can handle up to 300k input tokens, with comparable performance to ChatGPT and Claude 3.5 Sonnet.
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGoogle's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.
Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted Lite
Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-Derestricted-LiteQwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted Lite is a lighter-tuned community finetune for responsive multimodal chat, expressive writing, and roleplay.
Gemini Flash Lite Latest
google/gemini-flash-lite-latestCompatibility alias that routes to the newest version of Gemini Flash Lite. Currently routes to Gemini 3.5 Flash Lite.
ByteDance Seed 2.0 Lite
bytedance-seed/seed-2.0-liteByteDance Seed 2.0 Lite is a balanced long-context model for high-frequency enterprise workloads, tuned for unstructured information processing, text creation, search and recommendation, and stable structured outputs. Supports a 262k context window.