Ling 3.1 Flash is inclusionAI's hybrid reasoning model for coding, tool use, planning, and long-context agent workflows. It has 560B total parameters with 25B active parameters per token. Thinking is enabled by default and can be turned off in settings.
Added Sep 30, 2026
Context Window
262.1K
Max Output
32.8K
Input Price (Auto)
$0.075/1M
Output Price (Auto)
$0.22/1M
Cache Read (Auto)
$0.015/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Provider information for this model’s automatic routing. These routes cannot be selected individually.
Loading provider options…
Related text models
Compare Ling 3.1 Flash with similar models from the same provider or model family.
Ling 3.0 Flash VL
inclusionai/ling-3.0-flash-vlLing 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.
Ling 3.0 Flash
inclusionai/ling-3.0-flashLing-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.
Ling 3.0 Flash Thinking
inclusionai/ling-3.0-flash:thinkingLing-3.0-flash Thinking enables visible reasoning on inclusionAI's token-efficient 124B-parameter Mixture-of-Experts model for harder coding, tool use, planning, and production-scale agent workflows.
MiMo V2.6 Flash Uncensored Thinking
xiaomi/mimo-v2.6-flash-uncensored:thinkingMiMo V2.6 Flash Uncensored with maximum thinking enabled. Supports a 1M-token context window and tool calling.
MiMo V2.6 Flash Abliterated
xiaomi/mimo-v2.6-flash-abliteratedA MiMo V2.6 Flash finetune with refusal-direction ablation, image input, a 1M-token context window, separate reasoning output, and tool calling.
MiMo V2.6 Flash Uncensored
xiaomi/mimo-v2.6-flash-uncensoredA lower-refusal MiMo V2.6 Flash finetune with image input, a 1M-token context window, separate reasoning output, and tool calling.