GLM 5.3 Flash Uncensored

GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model with provider-dependent vision support, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.

  • Reasoning
  • Vision
  • Tool Calling
  • Structured Output

Added Jul 29, 2026

Model weights

Pricing

Auto routing · per 1M tokens
Input
$0.20
Output
$0.80
Cache read
$0.070
Compare provider prices

Specifications

Context window
1M
Max output
32.8K
Parameters
320B / 18B
Total / active
Avg output (7d)
748 tokens
Longer than 59% of models

Benchmarks

No public benchmark scores for this model yet.

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…