Qwen3.5 35B A3B Thinking

Qwen3.5 35B A3B with extended reasoning enabled. A native vision-language MoE model with hybrid attention.

  • Reasoning
  • Vision
  • Video Input

Added Feb 24, 2026

Model weights

Pricing

Auto routing · per 1M tokens
Input
$0.14
Output
$1.00
Cache read
$0.050
Compare provider prices

Specifications

Context window
260.1K
Max output
65.5K
Parameters
35B / 3B
Total / active
Avg output (7d)
1.8K tokens
Longer than 88% of models

Benchmarks

Sourced from Artificial Analysis.

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

89.2%

Better than 84% of models compared

Document reasoning

AA-LCR (unversioned / legacy)

Legacy fallback · Long context reasoning evaluation

72.0%

Better than 67% of models compared

Reasoning

HLE

Humanity's Last Exam

21.0%

Better than 69% of models compared

IFBench

Instruction-following benchmark

72.5%

Better than 89% of models compared

CritPt

Research-level physics reasoning

0.9%

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

26.5%

Better than 67% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

20.1%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

85.4%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

84.5%

Better than 77% of models compared

Last updated Sep 5, 2026, 12:00 AM

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…