Qwen3.8 Flash is Alibaba's latest fast multimodal model, with a million-token context window for coding, agentic workflows, visual understanding, long documents, codebases, and videos.
Added Aug 26, 2026
Model weightsContext Window
991.8K
Max Output
131.1K
Avg output tokens (7d)
854 tokens
Input Price (Auto)
$0.14/1M
Output Price (Auto)
$0.42/1M
Cache Read (Auto)
$0.016/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
39.9
Coding Index
73.1
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
55.9%
Better than 82% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1587 Elo
Better than 96% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1647 Elo
Better than 95% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
15.6%
Better than 64% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
79.7%
Better than 88% of models compared
Reasoning
HLE
Humanity's Last Exam
38.0%
Better than 89% of models compared
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
25.3%
Better than 84% of models compared
SciCode
Python programming for scientific computing
50.6%
Better than 61% of models compared
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
92.3%
Better than 94% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
79.7%
Better than 88% of models compared
Last updated Sep 11, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare Qwen3.8 Flash with similar models from the same provider or model family.
Qwen3.6 Flash
qwen/qwen3.6-flashQwen3.6 Flash is Alibaba's fast native vision-language model in the Qwen 3.6 family. It improves over 3.5 Flash with stronger coding/agent performance and better spatial intelligence, including object localization and detection.
Qwen3.8 Max 0902
qwen/qwen3.8-max-0902Qwen3.8 Max 0902 is Alibaba's September 2 checkpoint of its flagship Qwen3.8 Max model for coding, knowledge work, data analysis, and long-running agent workflows. It supports text, image, video, PDF input, selectable thinking, tool calling, structured output, and a near-million-token context window.
Qwen3.6 27B
qwen/qwen3.6-27bQwen3.6 27B is a native vision-language dense model with stronger agentic coding and STEM reasoning than Qwen 3.5 27B. It also improves spatial intelligence (including object localization/detection), plus video understanding, document OCR, and visual-agent workflows.
Qwen3.6 27B Thinking
qwen/qwen3.6-27b:thinkingQwen3.6 27B is a native vision-language dense model with stronger agentic coding and STEM reasoning than Qwen 3.5 27B. It also improves spatial intelligence (including object localization/detection), plus video understanding, document OCR, and visual-agent workflows.
Qwen3.7 Flash
qwen/qwen3.7-flashQwen3.7 Flash is Qwen's fast multimodal model for coding, search and computer-use agents, visual understanding, object recognition, spatial reasoning, and stable end-to-end task execution.
Qwen3.7 Flash Thinking
qwen/qwen3.7-flash:thinkingQwen3.7 Flash with thinking enabled for deeper multimodal reasoning, coding, search and computer-use agents, spatial reasoning, and multi-step task execution.