Provider logo

Qwen3 30B A3B Instruct 2507

qwen3-30b-a3b-instruct-2507
Provider logo

Qwen3 30B A3B Instruct 2507

qwen3-30b-a3b-instruct-2507

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. Significant improvements in general capabilities, including instruction following, logical reasoning, text comprehension, mathematics, science, coding and tool usage.

Added Feb 20, 2025

Model weights

Context Window

256.0K

Max Output

32.8K

Input Price (Auto)

$0.20/1M

Output Price (Auto)

$0.50/1M

Cache Read (Auto)

$0.10/1M

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

8.9

Better than 28% of models compared

Reasoning

GPQA Diamond

Graduate-level scientific reasoning

65.9%

Better than 43% of models compared

HLE

Humanity's Last Exam

6.9%

Better than 47% of models compared

IFBench

Instruction-following benchmark

33.1%

Better than 21% of models compared

T²-Bench Telecom

Conversational AI agents in dual-control scenarios

10.2%

Better than 7% of models compared

AA-LCR

Long context reasoning evaluation

25.3%

Better than 33% of models compared

CritPt

Research-level physics reasoning

0.0%

Coding

SciCode

Python programming for scientific computing

30.4%

Better than 43% of models compared

Terminal-Bench Hard

Agentic coding and terminal use

6.1%

Better than 34% of models compared

LiveCodeBench

Contamination-free coding benchmark

51.5%

Better than 58% of models compared

Math

AIME 2025

American Invitational Mathematics Examination 2025

66.3%

Better than 61% of models compared

AIME

American Invitational Mathematics Examination

72.7%

Better than 82% of models compared

Math-500

Diverse mathematical problem solving benchmark

97.5%

Better than 89% of models compared

Knowledge

MMLU-Pro

Professional and academic subject knowledge

77.7%

Better than 58% of models compared

AA-Omniscience Accuracy

Proportion of correctly answered questions

12.0%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

88.5%

Last updated Aug 16, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare Qwen3 30B A3B Instruct 2507 with similar models from the same provider or model family.

Qwen3 Coder 30B A3B Instruct

qwen3-coder-30b-a3b-instruct

Qwen3 Coder 30B with 3B active parameters, optimized for code generation and technical tasks

Qwen3 Next 80B A3B (Instruct)

Qwen/Qwen3-Next-80B-A3B-Instruct

Based on the new Qwen3‑Next architecture (hybrid attention, highly sparse MoE, training‑stability optimizations, and multi‑token prediction), the Qwen3‑Next‑80B‑A3B‑Instruct model delivers extreme efficiency with only 3B active parameters per pass. It performs comparably to Qwen3‑235B‑A22B‑Instruct‑2507 and shows clear advantages on ultra‑long context tasks (up to 256K tokens).

Qwen 3 235b A22B 2507

Qwen/Qwen3-235B-A22B-Instruct-2507

Qwen 3 235b A22B Instruct 2507 the updated version of Qwen3 235B A22B, with significant improvements in performance. This model is non-thinking.

Qwen3 30B A3B

qwen/qwen3-30b-a3b

Qwen 3 30b A3B is a 30b model with 3 billion active parameters per pass. Supports switching between thinking and non thinking: trigger thinking with /think and /no_think anywhere in a prompt or system message to toggle chain-of-thought reasoning.

Qwen 3.6 35B A3B Uncensored

qwen/qwen3.6-35b-a3b-uncensored

Qwen 3.6 35B A3B Uncensored is an NVFP4 open-weight mixture-of-experts model LoRA-tuned for fewer refusals across chat, coding, tool use, and multimodal tasks.

Qwen3.6 35B A3B TEE

TEE/qwen3.6-35b-a3b

Qwen3.6 35B A3B is an open-weight MoE model from Alibaba's Qwen team with 35B total parameters and 3B active parameters per token. Running inside a TEE (Trusted Execution Environment), with provider attestation support.