Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max for coding, complex analysis, and long-running agent workflows. It accepts text, image, and video input, supports tool calling and structured output, and has a 1M-token context window. Reasoning is always enabled.
Added Sep 23, 2026
Context Window
1.0M
Max Output
131.1K
Input Price (Auto)
$4.00/1M
Output Price (Auto)
$12.00/1M
Cache Read (Auto)
$0.50/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Qwen3.8 Max Prime with similar models from the same provider or model family.
Qwen3 Max
qwen/qwen3-maxQwen3 Max improves accuracy in coding and science, instruction following, and tool calling.
Qwen 3.8 27B Hemingway
qwen/qwen3.8-27b-hemmingwayQwen 3.8 27B Hemingway is an open-weight NVFP4 multimodal creative finetune for long-form prose, character dialogue, storytelling, and roleplay.
Qwen 3.8 27B Cybersecurity
qwen/qwen3.8-27b-cybersecurityQwen 3.8 27B Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports optional reasoning, image understanding, tool calling, and a 262,144-token context window.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an FP8 open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 524,288-token context window.
Qwen 3.8 27B Fable
qwen/qwen3.8-27b-fableQwen 3.8 27B Fable is an FP8 open-weight multimodal creative finetune for expressive dialogue, long-form storytelling, character work, and roleplay, with a 524,288-token context window.
Qwen 3.8 27B Obliterated
qwen/qwen3.8-27b-obliteratedQwen 3.8 27B Obliterated is an FP8 open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.