The first long-context LRM trained with reinforcement learning for long-context reasoning. Outperforms flagship models like o3-mini and achieves performance on par with Claude 3.7 Sonnet Thinking, demonstrating leading performance for long-context document QA tasks.
Context Window
128.0K
Max Output
41.0K
Input Price (Auto)
$0.14/1M
Output Price (Auto)
$0.60/1M
Cache Read (Auto)
$0.070/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare QwenLong L1 32B with similar models from the same provider or model family.
Qwen 2.5 Coder 32b
Qwen/Qwen2.5-Coder-32B-InstructThe latest series of Code-Specific Qwen large language models.
Qwen 2.5 32B Abliterated
huihui-ai/Qwen2.5-32B-Instruct-abliteratedUncensored version of Qwen 2.5 32B Instruct with restrictions removed.
Qwen 3 32b
qwen/qwen3-32bQwen 3 32b is a 32b model. Supports switching between thinking and non thinking: trigger thinking with /think and /no_think anywhere in a prompt or system message to toggle chain-of-thought reasoning.
Qwen 3.6 35B A3B Uncensored
qwen/qwen3.6-35b-a3b-uncensoredQwen 3.6 35B A3B Uncensored is an NVFP4 open-weight mixture-of-experts model LoRA-tuned for fewer refusals across chat, coding, tool use, and multimodal tasks.
Qwen 3.8 27B Uncensored
qwen/qwen3.8-27b-uncensoredQwen 3.8 27B Uncensored is an NVFP4 open-weight multimodal model LoRA-tuned for fewer refusals across chat, coding, tool use, and long-context work.
Qwen3.5 0.8B
qwen3.5-0.8bQwen3.5 0.8B is a lightweight open-weight multimodal model from Alibaba for fast reasoning, visual understanding, tool use, and JSON output.