Inkling

The non-thinking version of Thinking Machines' 975B-parameter open-weights Mixture-of-Experts generalist with 41B active parameters. It gives faster direct answers across text, images, and audio, and is built for agentic coding, tool use, detailed instruction following, and long-context work.

  • Vision
  • Audio Input
  • Tool Calling
  • Structured Output

Added Jul 15, 2026

Model weights

Pricing

Auto routing · per 1M tokens
Input
$0.95
Output
$4.05
Cache read
$0.16
Compare provider prices

Specifications

Context window
1M
Max output
32.8K
Parameters
975B / 41B
Total / active
Avg output (7d)
834 tokens
Longer than 63% of models

Benchmarks

Sourced from Artificial Analysis.

Intelligence Index

25.0

Better than 75% of models compared

Coding Index

52.1

Better than 62% of models compared

Agentic Index

22.5

Better than 64% of models compared

Agentic work

AutomationBench-AA

Workflow automation with guardrail penalties

5.0%

Better than 29% of models compared

AutomationBench-AA Tasks Completed

Fully completed workflows without guardrail violations

0.3%

Better than 3% of models compared

AA-Briefcase

Agentic knowledge work (Elo)

832 Elo

Better than 36% of models compared

GDPval-AA v2

Economically valuable tasks (Elo)

1079 Elo

Better than 52% of models compared

Document reasoning

GDP.pdf

Professional PDF reasoning: all-pass rate

12.8%

Better than 51% of models compared

AA-LCR v1.1

Long context reasoning with updated grading

77.3%

Better than 78% of models compared

MLCR-AA

Medical long-context reasoning

12.2%

Better than 38% of models compared

Reasoning

HLE

Humanity's Last Exam

31.9%

Better than 79% of models compared

CritPt

Research-level physics reasoning

5.4%

Coding

Terminal-Bench v4.0

Practical coding and terminal tasks

1.0%

Better than 42% of models compared

SciCode

Python programming for scientific computing

47.0%

Better than 45% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

41.5%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

67.7%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

87.2%

Better than 83% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

77.3%

Better than 78% of models compared

GDPval-AA (unversioned / legacy)

Economically valuable tasks

28.9%

Last updated Oct 5, 2026

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…