LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #150 OF 228 ASSISTANTS · OVERALL #240 OF 6,618 · VIBE SCORE 5.7 · 34 VOICES
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
2 mentions
1 mentions
26 weeks · 57 voices
HONEYMOON FADING
Early praise is cooling off in recent weeks.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Not enough discussion about Qwen3.5 397B A17B yet to call it. We found 34 posts but people did not say much either way.
29 POSTS · 5 COMMENTS · GITHUB · REDDIT · DEV FORUMS · BLUESKY · LEMMY · DEV.TO
Other models the crowd has fully reviewed, starting with multimodal models like this one.
Qwen 3.5
Qwen3.5-397B-A17B: First open-weight model in Qwen3.5 series released with benchmarks
New LLM Debate Benchmark: models debate the same motion twice with sides swapped in 10 turns. A wide variety of controversial and relevant topics. Sonnet 4.6 (high) wins. GLM-5 is the open weights leader.
All Qwen model oneshots: 1109 outputs to look at and compare!
Sonnet 4.6 scores on the Extended NYT Connections benchmark
Professional-grade local AI on consumer hardware — 80B stable on 44GB mixed VRAM (RTX 5060 Ti ×2 + RTX 3060) for under €800 total. Full compatibility matrix included.
When are we gonna get more 1-Bit models(Medium & Large size)?
Replaced $40/month in AI API subscriptions with self-hosted Ollama + n8n
XYZAILab/XYZ-Aquila-mini · Hugging Face
losing my mind fine-tuning jina-v5 for a legal corpus
Anyone successfully running Qwen3.5-397B-A17B-GPTQ-Int4?
To maximize prompt and generation throughput for Qwen3.5-397B-A17B-GPTQ-Int4 on 8xA6000, increase --max-num-batched-tokens (e.g., 16384 or higher), and ensure --gpu-memory-utilization is set high (e.g., 0.95) to maximize KV cache. Also, avoid duplicate…