LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #290 OF 324 ASSISTANTS · OVERALL #454 OF 7,062 · VIBE SCORE 5.1 · 25 VOICES
Huihui-Qwen3-VL-4B-Instruct-abliterated-GGUF on Hugging Face (image text to text). 8,947 downloads. Open weights for local or hosted use.
2 mentions
3 mentions
26 weeks · 23 voices
HOLDING UP
Recent vibe is steady. No clear fade in the crowd.
Open weights. You can use a hosted service, or download it and run it yourself, free.
People are mostly frustrated with Qwen3 VL 4B Instruct right now.
19 POSTS · 6 COMMENTS · REDDIT · GITHUB
1 thumbs up · 6 thumbs down
Other models the crowd has fully reviewed, starting with multimodal models like this one.
Ollama Models Ranked by VRAM Requirements
microsoft/Mage-VL · Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model
Krea 2 LoRA training on a 16GB RTX 5080: full measurements, and four sourced corrections to the guidance going around
Krea2 GGUF Worklfow for 8-12GB VRAM
Trying out LoKr instead of LoRA on Krea2
What do you use for image-to-text? This one doesn't seem to work
Write skip-module names vLLM can match after it fuses the MLP projections
Attempting to generate FLUX.2 Klein images from inside Canvas reports "no Z-Image model available"
Deployed FLUX.2 Klein LoRAs always report as "not deployed" in Canvas
perf(qwen_vl): multimodal RoPE leaks f32 into the decode residual
[BUG]: Agent mode (@agent / Agent chat mode) never returns a response with local Ollama provider — request aborted immediately (AbortError)
Describe the bug Running a Qwen3-VL-architecture GGUF VLM via runtime_id=llama_cpp with an image input produces different, backend-dependent results: - compute_unit=npu (HTP0/Hexagon): generation completes with no error, but the vision output is silently…