Browse all Inclusionai text models
Provider logo

Ling 3.1 Flash

inclusionai/ling-3.1-flash
Back
Provider logo

Ling 3.1 Flash

inclusionai/ling-3.1-flash
Back

Ling 3.1 Flash is inclusionAI's hybrid reasoning model for coding, tool use, planning, and long-context agent workflows. It has 560B total parameters with 25B active parameters per token. Thinking is enabled by default and can be turned off in settings.

Added Sep 30, 2026

Context Window

262.1K

Max Output

32.8K

Input Price (Auto)

$0.075/1M

Output Price (Auto)

$0.22/1M

Cache Read (Auto)

$0.015/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Provider information for this model’s automatic routing. These routes cannot be selected individually.

Loading provider options…

Compare Ling 3.1 Flash with similar models from the same provider or model family.

Ling 3.0 Flash VL

inclusionai/ling-3.0-flash-vl

Ling 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.

Ling 3.0 Flash

inclusionai/ling-3.0-flash

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.

Ling 3.0 Flash Thinking

inclusionai/ling-3.0-flash:thinking

Ling-3.0-flash Thinking enables visible reasoning on inclusionAI's token-efficient 124B-parameter Mixture-of-Experts model for harder coding, tool use, planning, and production-scale agent workflows.

MiMo V2.6 Flash Uncensored Thinking

xiaomi/mimo-v2.6-flash-uncensored:thinking

MiMo V2.6 Flash Uncensored with maximum thinking enabled. Supports a 1M-token context window and tool calling.

MiMo V2.6 Flash Abliterated

xiaomi/mimo-v2.6-flash-abliterated

A MiMo V2.6 Flash finetune with refusal-direction ablation, image input, a 1M-token context window, separate reasoning output, and tool calling.

MiMo V2.6 Flash Uncensored

xiaomi/mimo-v2.6-flash-uncensored

A lower-refusal MiMo V2.6 Flash finetune with image input, a 1M-token context window, separate reasoning output, and tool calling.