A fast, cost-effective reasoning model with 1M token context. Supports extended thinking with adjustable depth levels and built-in web grounding and code interpreter tools.
Context Window
1.0M
Max Output
65.5K
Input Price (Auto)
$0.51/1M
Output Price (Auto)
$4.25/1M
Cache Read (Auto)
$0.26/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
8.7
Agentic work
T²-Bench Telecom (legacy)
Legacy fallback · Conversational AI agents in dual-control scenarios
62.0%
Better than 57% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
18.7%
Better than 22% of models compared
Reasoning
HLE
Humanity's Last Exam
2.9%
Better than 2% of models compared
IFBench
Instruction-following benchmark
40.5%
Better than 40% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
Terminal-Bench Hard (legacy)
Legacy fallback · Agentic coding and terminal use
6.8%
Better than 37% of models compared
LiveCodeBench
Contamination-free coding benchmark
34.6%
Better than 42% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
33.7%
Better than 34% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
74.3%
Better than 46% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
14.0%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
85.9%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
60.3%
Better than 35% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
18.7%
Better than 22% of models compared
Last updated Sep 13, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Amazon Nova 2 Lite with similar models from the same provider or model family.
Amazon Nova Lite 1.0
amazon/nova-lite-v1Amazon's new lower cost model. Can handle up to 300k input tokens, with faster output but less thorough understanding than Amazon's Nova Pro.
Amazon Nova Pro 1.0
amazon/nova-pro-v1Amazon's new flagship model. Can handle up to 300k input tokens, with comparable performance to ChatGPT and Claude 3.5 Sonnet.
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-liteGoogle's cost-efficient Gemini 3.5 Flash Lite model for high-volume multimodal reasoning, tool use, and structured-output workloads.
Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted Lite
Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-Derestricted-LiteQwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted Lite is a lighter-tuned community finetune for responsive multimodal chat, expressive writing, and roleplay.
Gemini Flash Lite Latest
google/gemini-flash-lite-latestCompatibility alias that routes to the newest version of Gemini Flash Lite. Currently routes to Gemini 3.5 Flash Lite.
ByteDance Seed 2.0 Lite
bytedance-seed/seed-2.0-liteByteDance Seed 2.0 Lite is a balanced long-context model for high-frequency enterprise workloads, tuned for unstructured information processing, text creation, search and recommendation, and stable structured outputs. Supports a 262k context window.