LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #3 OF 12 CODE · OVERALL #152 OF 6,566 · VIBE SCORE 6.1 · 138 VOICES
Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...
30 mentions
12 mentions
26 weeks · 134 voices
QUIETLY DEGRADING
Recent vibe and chatter are both sliding.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Qwen3 Coder Next once and it's yours. No subscription, no rate limits, works offline.
80B total parameters with 3B activated per token (MoE architecture), a Q4 quantized file is about 45 GB to 48 GB
Nvidia GeForce RTX 3060 12GB plus 64 GB DDR4 system RAM
12 GB VRAM AND 64 GB RAM
~$220 USED GPU OR $650IQ3_XXS or Q4 GGUF with partial GPU offloading, about 10 to 18 tok/s at 16k context
Nvidia GeForce RTX 3090 24GB plus 64 GB DDR5 system RAM
24 GB VRAM AND 64 GB RAM
~$700 USED GPU ORQ4_K_M quant with hybrid GPU/CPU MoE offloading, ~30 tok/s at 64k to 120k context
Apple Mac Studio M2 Ultra with 128 GB Unified Memory
128 GB UNIFIED MEMORY
~$3,200 REFURBISHEDQ8_0 or full FP16 precision fully in memory, 35+ tok/s at full 256k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Qwen3 Coder Next takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNvidia GeForce RTX 3060 12GB plus 64 GB DDR4 system RAM | SWEET SPOTNvidia GeForce RTX 3090 24GB plus 64 GB DDR5 system RAM | FULL POWERApple Mac Studio M2 Ultra with 128 GB Unified Memory |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~36 S★★★★★ | ~17 S★★★★★ | —★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~1.8 MIN★★★★★ | ~50 S★★★★★ | —★★★★★ |
| Build a websitea small landing page, markup and styles together | ~9.5 MIN★★★★★ | ~4.4 MIN★★★★★ | —★★★★★ |
| Build a backendan API with routes, storage and tests | ~30 MIN★★★★★ | ~14 MIN★★★★★ | —★★★★★ |
| Build a gamea playable browser game in one file | ~18 MIN★★★★★ | ~8.3 MIN★★★★★ | —★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. A DASH MEANS NO SPEED WAS RESEARCHED FOR THAT RIG. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People are mostly frustrated with Qwen3 Coder Next right now. Biggest gripe: saying no too much.
102 POSTS · 36 COMMENTS · BLUESKY · HACKER NEWS · LEMMY · DEV FORUMS · REDDIT · GITHUB · DEV.TO · STACK OVERFLOW · BLOG · X
4 thumbs up · 4 thumbs down
0 thumbs up · 3 thumbs down
Other models the crowd has fully reviewed, starting with code models like this one.
Qwen3-Coder-Next
Qwen3-Coder-Next is released! 💜
I keep coming back to Qwen... Over and Over. Is there really nothing better under 120B?
Qwen3-Coder-Next is now the #1 most downloaded model on Unsloth!
Qwen3-Coder-Next GGUFs updated - now produces much better outputs!
Qwen3-Coder-Next GGUF Aider Coding Benchmarks
We created a Tool Calling Guide for LLMs!
Alibaba releases Qwen3-Coder-Next model with benchmarks
Free $200 AWS credits for new users — and yes, you can use them for AI models too. AWS Free Tier gives you 6 months to e
Qwen3.6 finally makes my Local LlaMa useful
Running Qwen3-Coder-Next-BF16 on 12GB VRAM
Best Local Model for Coding