LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #192 OF 228 ASSISTANTS · OVERALL #295 OF 6,618 · VIBE SCORE 5.4 · 82 VOICES
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
9 mentions
11 mentions
26 weeks · 134 voices
HOLDING UP
Recent vibe is steady. No clear fade in the crowd.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Qwen3.5 27B once and it's yours. No subscription, no rate limits, works offline.
27B parameters, a Q4 file is about 16.8 GB, full FP16 is about 55 GB
Desktop PC with 32 GB DDR5 RAM and RTX 3060 12GB
12 GB VRAM PLUS 32 GB RAM
~$280 USED GPU OR $650Q3_K or Q4_K with partial GPU offloading, ~6 tok/s, 8k context
Nvidia GeForce RTX 3090 24GB or RTX 4090 24GB
24 GB VRAM
~$700 USED RTX 3090Q4_K_M fully in VRAM, ~35 tok/s, 32k context
Dual Nvidia RTX 3090 24GB (48GB) or Mac Studio M2 Ultra (64G
48 GB TO 64 GB UNIFIED MEMORY OR VRAM
~$1400 USED DUAL GPUQ8_0 or full BF16 precision, ~45 tok/s, 128k+ context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Qwen3.5 27B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMDesktop PC with 32 GB DDR5 RAM and RTX 3060 12GB | SWEET SPOTNvidia GeForce RTX 3090 24GB or RTX 4090 24GB | FULL POWERDual Nvidia RTX 3090 24GB (48GB) or Mac Studio M2 Ultra (64G |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~83 S★★★★★ | ~14 S★★★★★ | ~11 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~4.2 MIN★★★★★ | ~43 S★★★★★ | ~33 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~22 MIN★★★★★ | ~3.8 MIN★★★★★ | ~3 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~69 MIN★★★★★ | ~12 MIN★★★★★ | ~9.3 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~42 MIN★★★★★ | ~7.1 MIN★★★★★ | ~5.6 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People are mostly frustrated with Qwen3.5 27B right now.
66 POSTS · 16 COMMENTS · HACKER NEWS · DEV FORUMS · GITHUB · REDDIT · LEMMY · STACK OVERFLOW · X
0 thumbs up · 7 thumbs down
Other models the crowd has fully reviewed, starting with multimodal models like this one.
microsoft/Fara1.5-27B · Hugging Face
We got 207 tok/s with Qwen3.5-27B on an RTX 3090
I ran Gemma 4 26B vs Qwen 3.5 27B across 18 real local business tests on my RTX 4090. Gemma won 13 to 5.
Reasoning-Medical0.1-27B (Qwen3.5-27B medical finetune, claims to surpass MedGemma)
Qwen 3.5 27B/35BA3B Tool Calling Issues: Why It Breaks & How I Fixed It
Qwen3.8 27B vs Qwen3.6 27B vs Qwen3.5 27B, a slight improvement in oneshotting ability across generations.
Qwen just released Qwen 3.5 medium model Series: Qwen 3.5 Flash plus 3 more models
Open-source models are closing the coding gap with GPT/Claude/Gemini ~1.5x faster than the frontier is advancing, and on decontaminated benchmarks a 27B model already beats Claude Opus 4.8 [live dashboard + analysis]
Deepseek V4 Flash 2-bit quant is the first model I can run locally that achieves 100% in this SQL benchmark
Qwen3.5 27B Uncensored Heretic Native MTP Preserved is Out Now With the Full 15 MTPs Preserved and Retained, Available in Safetensors, GGUFs, NVFP4, NVFP4 GGUFs and GPTQ-Int4 Formats!
Qwen3.5 27B Uncensored Heretic Native MTP Preserved is Out Now With the Full 15 MTPs Preserved and Retained, Available in Safetensors, GGUFs, NVFP4, NVFP4 GGUFs and GPTQ-Int4 Formats!
Why is Qwen3.5:27b using over 24GB of VRAM?