Private AI
Ling-3.0-flash Thinking enables visible reasoning on inclusionAI's token-efficient 124B-parameter Mixture-of-Experts model for harder coding, tool use, planning, and production-scale agent workflows.
Added Jul 23, 2026
Context Window
262.1K
Max Output
32.8K
Input Price (Auto)
$0.060/1M
Output Price (Auto)
$0.18/1M
Cache Read (Auto)
$0.012/1M
Capabilities
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…