LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #277 OF 296 ASSISTANTS · OVERALL #433 OF 6,736 · VIBE SCORE 5.0 · 25 VOICES
qwen3-moe-tiny on Hugging Face (text generation). 2,149 downloads. Open weights for local or hosted use.
0 mentions
1 mentions
26 weeks · 39 voices
HOLDING UP
Recent vibe is steady. No clear fade in the crowd.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Not enough discussion about Qwen3 Moe Tiny yet to call it. We found 25 posts but people did not say much either way.
24 POSTS · 1 COMMENTS · GITHUB · REDDIT
Other models the crowd has fully reviewed, starting with text models like this one.
fix(container/vllm): patch exaone_moe lm_head prefix + bump vLLM to v0.28.0 for K-EXAONE
feat(mlx): support Qwen3.8-Flash-Next and GLM-5.3-Flash — differentiable MoE routing, structural patch selection, and text-path loading
fix(container/vllm): patch exaone_moe lm_head prefix so K-EXAONE 2.0 NVFP4 loads
Align tiny/small Qwen2.5 config with Qwen/Qwen2.5-32B-Instruct
Fix qwen3_moe.py sanitize() gating and update DeepSeek-V3.2 to use PipelineMixin
fix(ENG-EXPERT-STREAM): the step clock had no caller, so the decode number measured a dead cache (#912, #1066)
fix(ENG-EXPERT-STREAM): the step clock had no caller, so the decode number measured a dead cache (#912, #1066)
fix(ENG-EXPERT-STREAM): the step clock had no caller, so the decode number measured a dead cache (#912, #1066)
[AMD][CI] Prune dead Qwen coverage from the AMD nightly
fix: keep torch cos/sin for the compile-time decoderonly rotary constants
[models-dsv2lite-moe] DeepSeek-Coder-V2-Lite: greedy/ungrouped MoE router
feat(models): port DeepSeek-Coder-V2-Lite's real greedy MoE router (issue #218)