Browse all PrismML text models

Ternary Bonsai 2 27B

prism-ml/ternary-bonsai-2-27b

Ternary Bonsai 2 27B

prism-ml/ternary-bonsai-2-27b

Ternary Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window.

Added Sep 18, 2026

Model weights

Context Window

262.1K

Max Output

32.8K

Input Price (Auto)

$0.075/1M

Output Price (Auto)

$0.50/1M

Cache Read (Auto)

$0.037/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Ternary Bonsai 2 27B with similar models from the same provider or model family.

Qwen 3.8 27B Queen

qwen/qwen3.8-27b-queen

Qwen 3.8 27B Queen is an FP8 open-weight roleplay finetune with image understanding, tool calling, optional reasoning, and a 524,288-token context window.

Qwen3.8 27B TEE

TEE/qwen3.8-27b

Qwen3.8 27B is an open-weight dense vision-language model from Alibaba for reasoning, coding, professional workflows, multimodal interaction, tool use, and structured output. Running inside a TEE (Trusted Execution Environment), with provider attestation support.

Qwen3.8 27B

qwen/qwen3.8-27b

Qwen3.8 27B is an open-weight multimodal model from Alibaba for coding, visual understanding, tool use, and structured output. This variant keeps thinking disabled for faster direct responses. Requests containing video cost $0.32 per million input tokens and $2.50 per million output tokens.

Qwen3.8 27B Thinking

qwen/qwen3.8-27b:thinking

Qwen3.8 27B is an open-weight multimodal model from Alibaba for reasoning, coding, visual understanding, tool use, and structured output. This variant enables thinking by default. Requests containing video cost $0.32 per million input tokens and $2.50 per million output tokens.

Qwen 3.8 27B Fable

qwen/qwen3.8-27b-fable

Qwen 3.8 27B Fable is an FP8 open-weight multimodal creative finetune for expressive dialogue, long-form storytelling, character work, and roleplay, with a 524,288-token context window.

Qwen 3.8 27B Obliterated

qwen/qwen3.8-27b-obliterated

Qwen 3.8 27B Obliterated is an FP8 open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, reasoning, tool use, and long-context work.