Qwen3.5 35B A3B

Qwen3.5 35B A3B is a native vision-language MoE model with hybrid attention designed for efficient inference and strong general performance.

  • Vision
  • Video Input

Added Feb 24, 2026

Model weights

Pricing

Auto routing · per 1M tokens
Input
$0.14
Output
$1.00
Cache read
$0.050
Compare provider prices

Specifications

Context window
260.1K
Max output
65.5K
Parameters
35B / 3B
Total / active
Avg output (7d)
1.1K tokens
Longer than 71% of models

Benchmarks

Sourced from Artificial Analysis.

Coding Index

37.0

Better than 43% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

86.3%

Better than 79% of models compared

GDPval-AA (unversioned / legacy)

Legacy fallback · Economically valuable tasks

4.7%

Document reasoning

AA-LCR (unversioned / legacy)

Legacy fallback · Long context reasoning evaluation

63.0%

Better than 54% of models compared

Reasoning

HLE

Humanity's Last Exam

13.4%

Better than 59% of models compared

IFBench

Instruction-following benchmark

44.5%

Better than 50% of models compared

CritPt

Research-level physics reasoning

0.6%

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

10.6%

Better than 44% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

15.7%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

93.4%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

81.9%

Better than 70% of models compared

Last updated Sep 20, 2026, 12:01 AM

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…