Provider logo

Qwen 3 235b A22B 2507 Thinking

Qwen/Qwen3-235B-A22B-Thinking-2507
Provider logo

Qwen 3 235b A22B 2507 Thinking

Qwen/Qwen3-235B-A22B-Thinking-2507

The thinking version of Qwen 3 235b A22B 2507, with enhanced reasoning capabilities and step-by-step problem solving.

Added Sep 11, 2025

Model weights

Context Window

256.0K

Max Output

262.1K

Input Price (Auto)

$0.23/1M

Output Price (Auto)

$2.30/1M

Cache Read (Auto)

$0.20/1M

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

19.9

Better than 55% of models compared

Coding Index

22.1

Better than 29% of models compared

Agentic Index

3.8

Better than 23% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

79.0%

Better than 68% of models compared

HLE

Humanity's Last Exam

15.9%

Better than 70% of models compared

IFBench

Instruction-following benchmark

51.2%

Better than 61% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

53.2%

Better than 54% of models compared

AA-LCR

Long context reasoning evaluation

70.7%

Better than 81% of models compared

GDPval-AA

Economically valuable tasks

2.1%

CritPt

Research-level physics reasoning

0.0%

Coding

SciCode

Python programming for scientific computing

42.4%

Better than 78% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

13.6%

Better than 50% of models compared

LiveCodeBench

Contamination-free coding benchmark

78.8%

Better than 90% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

91.0%

Better than 92% of models compared

AIME

American Invitational Mathematics Examination

94.0%

Better than 98% of models compared

Math-500

Diverse mathematical problem solving benchmark

98.4%

Better than 95% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

84.3%

Better than 89% of models compared

AA-Omniscience Accuracy

Proportion of correctly answered questions

22.8%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

89.8%

Last updated Aug 16, 2026

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…

Compare Qwen 3 235b A22B 2507 Thinking with similar models from the same provider or model family.