LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #15 OF 268 ASSISTANTS · OVERALL #25 OF 6,566 · VIBE SCORE 7.1 · 1,004 VOICES
A fast, open AI model from Alibaba built to help with coding, writing, and general tasks.
95 mentions
27 mentions
26 weeks · 1,712 voices
QUIETLY DEGRADING
Recent vibe and chatter are both sliding.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Qwen3.8 2.4T A95B once and it's yours. No subscription, no rate limits, works offline.
2.4T parameters (95B active MoE), a 1-bit GGUF is about 397 GB
Mac Studio M3 Ultra (512GB Unified Memory)
512 GB UNIFIED MEMORY
~$9,500 USEDToo big for standard consumer hardware; runs 1-bit GGUF at 1 to 2 tok/s, extremely slow but fits entirely in memory
8x RTX 3090 24GB GPUs (custom workstation)
192 GB VRAM + 256 GB SYSTEM RAM
~$7,500 USEDDynamic 1-bit Standard (UD-IQ1_S) at 3 to 5 tok/s with partial GPU offloading
8x NVIDIA H200 141GB GPUs (server node)
1.1 TB VRAM
~$320,000 NEWFP8 or Q8_0 precision at 40 plus tok/s, full 262k context window
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Qwen3.8 2.4T A95B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMMac Studio M3 Ultra (512GB Unified Memory) | SWEET SPOT8x RTX 3090 24GB GPUs (custom workstation) | FULL POWER8x NVIDIA H200 141GB GPUs (server node) |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~5.6 MIN★★★★★ | ~2.1 MIN★★★★★ | —★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~17 MIN★★★★★ | ~6.3 MIN★★★★★ | —★★★★★ |
| Build a websitea small landing page, markup and styles together | ~89 MIN★★★★★ | ~33 MIN★★★★★ | —★★★★★ |
| Build a backendan API with routes, storage and tests | ~4.6 HR★★★★★ | ~1.7 HR★★★★★ | —★★★★★ |
| Build a gamea playable browser game in one file | ~2.8 HR★★★★★ | ~63 MIN★★★★★ | —★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. A DASH MEANS NO SPEED WAS RESEARCHED FOR THAT RIG. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Users love Qwen3.8 2.4T A95B for its fast, accurate answers and strong coding assistance. However, many complain about high costs, frequent outages, and a tendency to lecture or refuse prompts.
688 POSTS · 316 COMMENTS · BLUESKY · REDDIT · HACKER NEWS · X · LEMMY · GITHUB · DEV.TO · BLOG · DEV FORUMS
24 thumbs up · 4 thumbs down
8 thumbs up · 7 thumbs down
9 thumbs up · 2 thumbs down
5 thumbs up · 5 thumbs down
5 thumbs up · 1 thumbs down
0 thumbs up · 5 thumbs down
Checked picks first: tags say why you would switch, stars say how fully each one stands in for Qwen3.8 2.4T A95B.
We promised open weights for Qwen3.8. Now, time to meet them! 🎉 ⚡ Qwen3.8-27B: - A native multimodal dense model. With
Qwen-3.8 27b released! Significant jump! Qwen just released Qwen3.8-27B, a compact open-weight multimodal model that rep
Qwen3.8-2.4T-A95B Released
Qwen3.8-Max: A New Bar for Coding and Cowork
Qwen3.8-2.4T-A95B is out now!
Qwen3.8-2.4T-A95B (aka Qwen3.8-Max) open release time: next wednesday
Qwen3.8 Max now ranked as the best overall model by agentic index
A little appetizer before Qwen3.8-27B The Qwen and Qwen Cloud teams are going live this Friday at 10:00 AM (UTC+8), whic
Fable 5 refuses to touch Qwen deployments?
Hard to believe DeepSeek-V4-Pro-0813, Grok-4.6, and the Qwen3.8-Max weight release all happened within a single hour. Ev
DeepSeek V4 Pro 0813 scores 53 on the Artificial Analysis Intelligence Index, 8 points above April's DeepSeek V4 Pro - b
How do you plan to run Qwen3.8-2.4T-A95B locally?