Qwen3.8 Flash
Qwen3.8 Flash is Alibaba's latest fast multimodal model, with a million-token context window for coding, agentic workflows, visual understanding, long documents, codebases, and videos.
- Reasoning
- Vision
- Video Input
- Tool Calling
- Structured Output
Added Aug 26, 2026
Model weightsPricing
Auto routing · per 1M tokens- Input
- $0.14
- Output
- $0.42
- Cache read
- $0.016
Specifications
- Context window
- 991.8K
- Max output
- 131.1K
- Avg output (7d)
- 770 tokens
- Longer than 58% of models
Benchmarks
Benchmarks
Sourced from Artificial Analysis.
Intelligence Index
39.8
Coding Index
73.1
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
55.9%
Better than 77% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1588 Elo
Better than 93% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1612 Elo
Better than 93% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
15.6%
Better than 58% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
79.7%
Better than 85% of models compared
Reasoning
HLE
Humanity's Last Exam
38.0%
Better than 86% of models compared
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
25.3%
Better than 78% of models compared
SciCode
Python programming for scientific computing
50.6%
Better than 57% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
92.3%
Better than 94% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
79.7%
Better than 85% of models compared
Last updated Oct 2, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…