Qwen3-Coder

Runs on your device

Alibaba's Qwen team · Jul 2025

Alibaba's agentic coding model — a 30B MoE with only 3B active per token. Tuned for repository-scale work and tool use (Qwen Code, Cline).

Context
256K tokens
Smallest download
18.6 GB
Runs from
32 GB RAM
License
Apache 2.0
codingagenticlong contextMLX for Apple Silicon

Will Qwen3-Coder run on your machine?

This is the same sizing the app uses - not a marketing estimate.

30B-A3B (MoE)18.6 GB downloadRuns - experts in RAM, the rest on the GPU

Mixture-of-experts: the always-on layers sit on your card, the experts sit in system memory - bigger than your card, still quick.

On Apple Silicon, system memory is also graphics memory - pick "None / integrated / Apple Silicon" and read the bars against your Mac's memory. The app checks your real hardware before every download.

Variants & sizes

VariantDownloadQuantizationNeedsApple Silicon
30B-A3B (MoE)MoE18.6 GBQ4_K_M32 GB RAMMLX · 17.2 GB

The app picks the right variant for your machine and verifies every download. Model files come straight from their official repositories, pinned to exact revisions.

License

Qwen3-Coder is released under the Apache 2.0 - read the terms. The weights download to your machine and stay there; Your Own AI adds no accounts, tracking, or lock-in on top.

Frequently asked questions

Can I run Qwen3-Coder offline?
Yes. Qwen3-Coder is a downloadable model: with Your Own AI it runs entirely on your computer, works with no internet connection, and your conversations never leave your device.
What hardware does Qwen3-Coder need?
The smallest variant (30B-A3B (MoE), a 18.6 GB download) runs on a machine with 32 GB of memory. Larger variants want more memory or a graphics card - the interactive checker on this page grades every variant against your hardware, using the same logic as the app.
Is Qwen3-Coder free to use?
The model weights are released under the Apache 2.0 license and download at no charge. Your Own AI itself is free and open source for local use.
How much context does Qwen3-Coder support?
256K tokens of context.

Run Qwen3-Coder in two clicks

Download Your Own AI, pick Qwen3-Coder, and the app sizes it to your hardware. Private by architecture - no account needed for local use.