Ternary Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window.
Added Sep 18, 2026
Model weightsContext Window
262.1K
Max Output
32.8K
Input Price (Auto)
$0.075/1M
Output Price (Auto)
$0.50/1M
Cache Read (Auto)
$0.037/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Ternary Bonsai 2 27B with similar models from the same provider or model family.
Qwen 3.8 27B Queen
qwen/qwen3.8-27b-queenQwen 3.8 27B Queen is an FP8 open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 524,288-token context window.
Qwen3.8 27B TEE
TEE/qwen3.8-27bQwen3.8 27B is an open-weight dense vision-language model from Alibaba for reasoning, coding, professional workflows, multimodal interaction, tool use, and structured output. Running inside a TEE (Trusted Execution Environment), with provider attestation support.
Qwen3.8 27B
qwen/qwen3.8-27bQwen3.8 27B is an open-weight multimodal model from Alibaba for coding, visual understanding, tool use, and structured output. This variant keeps thinking disabled for faster direct responses. Requests containing video cost $0.32 per million input tokens and $2.50 per million output tokens.
Qwen3.8 27B Thinking
qwen/qwen3.8-27b:thinkingQwen3.8 27B is an open-weight multimodal model from Alibaba for reasoning, coding, visual understanding, tool use, and structured output. This variant enables thinking by default. Requests containing video cost $0.32 per million input tokens and $2.50 per million output tokens.
Qwen 3.8 27B Fable
qwen/qwen3.8-27b-fableQwen 3.8 27B Fable is an FP8 open-weight multimodal creative finetune for expressive dialogue, long-form storytelling, character work, and roleplay, with a 524,288-token context window.
Qwen 3.8 27B Obliterated
qwen/qwen3.8-27b-obliteratedQwen 3.8 27B Obliterated is an FP8 open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.