Qwen: Qwen3 30B A3B

by Qwen

Qwen3, the latest generation in the Qwen large language model series, features both dense and mixture-of-experts (MoE) architectures to excel in reasoning, multilingual support, and advanced agent tasks. Its unique ability to switch seamlessly between a thinking mode for complex reasoning and a non-thinking mode for efficient dialogue ensures versatile, high-quality performance. Significantly outperforming prior models like QwQ and Qwen2.5, Qwen3 delivers superior mathematics, coding, commonsense reasoning, creative writing, and interactive dialogue capabilities. The Qwen3-30B-A3B variant includes 30.5 billion parameters (3.3 billion activated), 48 layers, 128 experts (8 activated per task), and supports up to 131K token contexts with YaRN, setting a new standard among open-source models.

Avg Score

57.7%

11 answers

Avg Latency

149.8s

8 runs

Pricing

$0.06

input

$0.22

output

per 1M tokens

Context

41K

tokens

Alternatives

Models with similar or better quality but different tradeoffs

No alternatives found

Run benchmarks on this model to discover alternatives

Other Models from Qwen

Compare performance with other models from the same creator

Model	Score	Latency	Cost/1M
Qwen: Qwen3 Coder 480B A35B	100.0%	1.1s	Free
Qwen: Qwen3 VL 235B A22B Instruct	88.6%	48.0s	$0.70
Qwen: Qwen3 Max	87.9%	31.7s	$3.60
Qwen: Qwen-Plus	85.8%	22.0s	$0.80
Qwen: Qwen3 235B A22B Thinking 2507	84.6%	248.4s	$0.35
Qwen: Qwen3 Coder Plus	84.2%	20.5s	$3.00
Qwen: Qwen3 Coder 480B A35B	83.5%	19.2s	$0.58
Qwen: Qwen3 Coder 480B A35B	81.3%	6.2s	$1.01
Qwen: Qwen3 Next 80B A3B Thinking	80.8%	28.8s	$0.68
Qwen: Qwen3 VL 235B A22B Thinking	80.4%	112.0s	$1.98
Qwen: Qwen3 VL 32B Instruct	78.8%	26.3s	$1.00
Qwen: Qwen Plus 0728	78.5%	19.6s	$0.80
Qwen: Qwen3 235B A22B	74.6%	78.3s	$0.40
Qwen: Qwen3 30B A3B Instruct 2507	73.1%	30.8s	$0.20
Qwen: Qwen3 Next 80B A3B Instruct	71.4%	27.9s	$0.59
Qwen: Qwen3 VL 8B Thinking	70.8%	79.4s	$1.14
Qwen: Qwen-Turbo	70.7%	18.6s	$0.13
Qwen: Qwen Plus 0728	70.0%	49.2s	$2.20
Qwen: Qwen3 235B A22B Instruct 2507	69.6%	25.6s	$0.27
Qwen: Qwen3 Coder Flash	69.2%	14.1s	$0.90
Qwen: Qwen3 30B A3B Thinking 2507	65.8%	45.6s	$0.20
Qwen: Qwen-Max	65.8%	12.7s	$4.00
Qwen: Qwen2.5 VL 72B Instruct	65.0%	22.1s	$0.38
Qwen: Qwen3 Coder 30B A3B Instruct	63.6%	30.8s	$0.17
Qwen: Qwen3 VL 30B A3B Thinking	63.1%	83.6s	$0.60
Qwen2.5 72B Instruct	62.9%	22.2s	$0.26
Qwen: Qwen3 8B	62.5%	111.5s	$0.15
Qwen: Qwen3 14B	61.5%	70.4s	$0.14
Qwen: Qwen VL Max	61.3%	36.4s	$2.00
Qwen: Qwen3 32B	59.6%	121.7s	$0.16
Qwen: QwQ 32B	59.2%	143.6s	$0.28
Qwen: Qwen VL Plus	57.9%	10.4s	$0.42
Qwen: Qwen3 VL 30B A3B Instruct	57.5%	38.3s	$0.38
Qwen: Qwen3 VL 8B Instruct	54.6%	133.7s	$0.29
Qwen2.5 Coder 32B Instruct	52.7%	19.7s	$0.07
Qwen: Qwen2.5 VL 32B Instruct	47.3%	39.1s	$0.14
Qwen: Qwen2.5-VL 7B Instruct	31.3%	20.7s	Free
Qwen: Qwen2.5 7B Instruct	30.8%	12.6s	$0.07
Qwen: Qwen2.5-VL 7B Instruct	21.7%	10.9s	$0.20
Qwen: Qwen2.5 Coder 7B Instruct	6.7%	6.4s	$0.06
Qwen: Qwen3 Next 80B A3B Instruct	—	—	Free
Qwen: Qwen3 4B (free)	—	—	Free

Benchmark Performance

How this model performs across different benchmarks

No benchmark data available

Run benchmarks with this model to see performance breakdown

Price vs Performance

Compare cost efficiency across all models

Current model (baseline)

Other models (relative score)

Y-axis shows score difference from shared benchmarks. X-axis uses log scale.

Score Over Time

Performance trends across all benchmark runs

Benchmark Activity

Number of benchmark runs over time

Quickstart

Get started with this model using OpenRouter

View on OpenRouter

import { OpenRouter } from "@openrouter/sdk";

const openrouter = new OpenRouter({
  apiKey: "<OPENROUTER_API_KEY>"
});

const completion = await openrouter.chat.completions.create({
  model: "qwen/qwen3-30b-a3b",
  messages: [
    {
      role: "user",
      content: "Hello!"
    }
  ]
});

console.log(completion.choices[0].message.content);

Get your API key at openrouter.ai/keys