qwen
30B parameters
Commercial OK
Reviewed July 2026

Qwen3 Coder 30B-A3B

Qwen3 Coder 30B-A3B is Alibaba's coding-specialized MoE (30B total, 3B active) under Apache-2.0, sitting at 7.4M Ollama pulls. It is the entry point into Alibaba's dedicated Qwen3-Coder line, whose smallest official variant starts at 30B.

License: Apache-2.0·Context: 131,072 tokens

Overview

Qwen3 Coder 30B-A3B is Alibaba's coding-specialized MoE (30B total, 3B active) under Apache-2.0, sitting at 7.4M Ollama pulls. It is the entry point into Alibaba's dedicated Qwen3-Coder line, whose smallest official variant starts at 30B.

Strengths

  • Apache-2.0 license with commercial use allowed
  • MoE (3B active) decodes fast relative to its 30B total size
  • High adoption: 7.4M Ollama pulls for the coding-specialist line

Weaknesses

  • No smaller official Qwen3-Coder variant exists below 30B
  • Coding-specialized training — general chat quality not the design target

Quantization variants

Each quantization trades model quality for file size and VRAM. Q4_K_M is the most popular starting point.

QuantizationFile sizeVRAM required
Q4_K_M19.0 GB23 GB

Get the model

Ollama

One-line install

ollama run qwen3-coder:30bRead our Ollama review →

Hardware that runs this

Cards with enough VRAM for at least one quantization of Qwen3 Coder 30B-A3B.

Compare alternatives

Models worth comparing

Same parameter band, plus what's one tier above and below — so you can decide what actually fits your hardware.

Frequently asked

What's the minimum VRAM to run Qwen3 Coder 30B-A3B?

23GB of VRAM is enough to run Qwen3 Coder 30B-A3B at the Q4_K_M quantization (file size 19.0 GB). Higher-quality quantizations need more.

Can I use Qwen3 Coder 30B-A3B commercially?

Yes — Qwen3 Coder 30B-A3B ships under the Apache-2.0, which permits commercial use. Always read the license text before deployment.

What's the context length of Qwen3 Coder 30B-A3B?

Qwen3 Coder 30B-A3B supports a context window of 131,072 tokens (about 131K).

How do I install Qwen3 Coder 30B-A3B with Ollama?

Run `ollama pull qwen3-coder:30b` to download, then `ollama run qwen3-coder:30b` to start a chat session. The default quantization is Q4_K_M.

Source: Vendor official documentation

Reviewed by RunLocalAI Editorial. See our editorial policy for how we research and verify model claims.

Related — keep moving

Before you buy

Verify Qwen3 Coder 30B-A3B runs on your specific hardware before committing money.