GLM 5.3 Flash Uncensored
GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model with provider-dependent vision support, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.
- Reasoning
- Vision
- Tool Calling
- Structured Output
Added Jul 29, 2026
Model weightsPricing
Auto routing · per 1M tokens- Input
- $0.20
- Output
- $0.80
- Cache read
- $0.070
Specifications
- Context window
- 1M
- Max output
- 32.8K
- Parameters
- 320B / 18B
- Total / active
- Avg output (7d)
- 748 tokens
- Longer than 59% of models
Benchmarks
No public benchmark scores for this model yet.
Providers
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…