Private AI
Ling-3.0-flash Thinking enables visible reasoning on inclusionAI's token-efficient 124B-parameter Mixture-of-Experts model for harder coding, tool use, planning, and production-scale agent workflows.
Added Jul 23, 2026
Model weightsContext Window
262.1K
Max Output
32.8K
Input Price (Auto)
$0.075/1M
Output Price (Auto)
$0.22/1M
Cache Read (Auto)
$0.015/1M
Capabilities
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
37.8
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Coding Index
50.6
Agentic Index
29.3
GPQA Diamond
Graduate-level scientific reasoning
85.5%
Better than 82% of models compared
HLE
Humanity's Last Exam
23.7%
Better than 80% of models compared
AA-LCR
Long context reasoning evaluation
67.0%
Better than 74% of models compared
GDPval-AA
Economically valuable tasks
30.4%
CritPt
Research-level physics reasoning
1.7%
SciCode
Python programming for scientific computing
41.1%
Better than 76% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
18.2%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
44.1%
Last updated Aug 13, 2026
Artificial Analysis