LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #1 OF 8 CODE · OVERALL #56 OF 6,618 · VIBE SCORE 6.6 · 415 VOICES
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
35 mentions
13 mentions
26 weeks · 451 voices
QUIETLY DEGRADING
Recent vibe and chatter are both sliding.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Qwen3 Coder 480B A35B once and it's yours. No subscription, no rate limits, works offline.
480B total parameters (35B active MoE), a Q4_K_M quant is about 280 GB while FP16 is around 960 GB
Apple Mac Studio M2 Ultra (192GB Unified Memory) + SSD swap
192 GB UNIFIED MEMORY
~$5,000 USEDModel is too massive for consumer GPUs; runs heavily quantized Q2/Q3 with partial SSD paging at sub-5 tok/s, limited con
Workstation with 6x NVIDIA RTX 3090 (24GB) via PCIe risers
144 GB VRAM + 256 GB SYSTEM DDR5 RAM
~$6,200 USEDQ4_K_M split across GPUs and system RAM via llama.cpp, ~8 to 12 tok/s, 16k context
Custom multi-GPU rig with 8x NVIDIA RTX 4090 (24GB) or 4x RT
192 GB TO 384 GB VRAM
~$16,000 TO $28,000Q4 or Q8 quant fully resident in VRAM, 25 tok/s, full 256k context window
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Qwen3 Coder 480B A35B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMApple Mac Studio M2 Ultra (192GB Unified Memory) + SSD swap | SWEET SPOTWorkstation with 6x NVIDIA RTX 3090 (24GB) via PCIe risers | FULL POWERCustom multi-GPU rig with 8x NVIDIA RTX 4090 (24GB) or 4x RT |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~1.7 MIN★★★★★ | ~50 S★★★★★ | ~20 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~5 MIN★★★★★ | ~2.5 MIN★★★★★ | ~60 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~27 MIN★★★★★ | ~13 MIN★★★★★ | ~5.3 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~83 MIN★★★★★ | ~42 MIN★★★★★ | ~17 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~50 MIN★★★★★ | ~25 MIN★★★★★ | ~10 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People are mostly happy with Qwen3 Coder 480B A35B. Biggest praise: help with code. Biggest gripe: working when you need it.
364 POSTS · 51 COMMENTS · BLUESKY · GITHUB · REDDIT · LEMMY · X · BLOG · DEV FORUMS
7 thumbs up · 3 thumbs down
3 thumbs up · 0 thumbs down
3 thumbs up · 0 thumbs down
2 thumbs up · 1 thumbs down
1 thumbs up · 2 thumbs down
Checked picks first: tags say why you would switch, stars say how fully each one stands in for Qwen3 Coder 480B A35B.
30 Days of an LLM Honeypot
China is winning the AI race for coding while being open source
Claude Code Competitor Just Dropped and it’s Open Source
Ollama Models Ranked by VRAM Requirements
Alibaba releases Qwen3-Coder
qwen3-coder is here
Alibaba releases Qwen3-Coder-Next model with benchmarks
Qwen3 Coder (free) is now available on OpenRouter. Go nuts.
I ran Gemma 4 26B vs Qwen 3.5 27B across 18 real local business tests on my RTX 4090. Gemma won 13 to 5.
You're paying $20/month for a coding agent that runs one model. Here's a free one that runs any of them. The coding-agen
GPT-5 releases in <15 hours. How do you think it will compare to Claude Opus?
Qwen 3.5 27B/35BA3B Tool Calling Issues: Why It Breaks & How I Fixed It