Qwen 3 8B is a 8B model. Supports switching between thinking and non thinking: trigger thinking with /think and /no_think anywhere in a prompt or system message to toggle chain-of-thought reasoning.
Context Window
41.0K
Max Output
32.8K
Avg output tokens (7d)
164 tokens
Input Price (Auto)
$0.47/1M
Output Price (Auto)
$0.47/1M
Cache Read (Auto)
$0.23/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
6.0
Agentic work
T²-Bench Telecom (legacy)
Legacy fallback · Conversational AI agents in dual-control scenarios
24.9%
Better than 28% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
0.0%
Better than 6% of models compared
Reasoning
HLE
Humanity's Last Exam
1.9%
Better than 0% of models compared
IFBench
Instruction-following benchmark
28.6%
Better than 12% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
Terminal-Bench Hard (legacy)
Legacy fallback · Agentic coding and terminal use
2.3%
Better than 20% of models compared
LiveCodeBench
Contamination-free coding benchmark
20.2%
Better than 22% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
24.3%
Better than 26% of models compared
AIME
American Invitational Mathematics Examination
24.3%
Better than 51% of models compared
Math-500
Diverse mathematical problem solving benchmark
82.8%
Better than 49% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
64.3%
Better than 26% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
11.1%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
95.6%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
45.2%
Better than 19% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
0.0%
Better than 6% of models compared
Last updated Sep 11, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Qwen 3 8B with similar models from the same provider or model family.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 262,144-token context window.
Qwen 3.8 27B Fable
qwen/qwen3.8-27b-fableQwen 3.8 27B Fable is an open-weight multimodal creative finetune for expressive dialogue, long-form storytelling, character work, and roleplay.
Qwen 3.8 27B Obliterated
qwen/qwen3.8-27b-obliteratedQwen 3.8 27B Obliterated is an open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.
Qwen 3.8 27B Obliterated Thinking
qwen/qwen3.8-27b-obliterated:thinkingQwen 3.8 27B Obliterated with thinking enabled for more deliberate creative work, coding, multimodal analysis, tool use, and long-context problem solving.
Qwen 3.8 27B Uncensored
qwen/qwen3.8-27b-uncensoredQwen 3.8 27B Uncensored is an NVFP4 open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.
Qwen 3.8 27B Uncensored Thinking
qwen/qwen3.8-27b-uncensored:thinkingQwen 3.8 27B Uncensored with thinking enabled for more deliberate creative work, coding, multimodal analysis, tool use, and long-context problem solving.