LOADING DATASHEET
LOADING DATASHEET
by InclusionAI
RANKED #117 OF 244 ASSISTANTS · OVERALL #199 OF 6,446 · VIBE SCORE 5.9 · 56 VOICES
Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native
20 mentions
1 mentions
26 weeks · 127 voices
AGING WELL
The crowd is warmer now than it was at the start.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Ling 3.0 tiny once and it's yours. No subscription, no rate limits, works offline.
7.9B total parameters (1.3B active per token MoE), a Q4/INT4 file is about 4.8 GB and FP8 is about 8.3 GB
NVIDIA GeForce RTX 3060 12GB
12 GB VRAM
~$250 USEDINT4/Q4 quantization, ~65 tok/s, 8k context
NVIDIA GeForce RTX 4070 12GB
12 GB VRAM
~$520 NEWFP8 quantization, ~95 tok/s, 16k context
NVIDIA GeForce RTX 4090 24GB
24 GB VRAM
~$1750 USEDBF16 unquantized, ~130 tok/s, 64k+ context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Ling 3.0 tiny takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3060 12GB | SWEET SPOTNVIDIA GeForce RTX 4070 12GB | FULL POWERNVIDIA GeForce RTX 4090 24GB |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~8 S★★★★★ | ~5 S★★★★★ | ~4 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~23 S★★★★★ | ~16 S★★★★★ | ~12 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~2.1 MIN★★★★★ | ~84 S★★★★★ | ~62 S★★★★★ |
| Build a backendan API with routes, storage and tests | ~6.4 MIN★★★★★ | ~4.4 MIN★★★★★ | ~3.2 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~3.8 MIN★★★★★ | ~2.6 MIN★★★★★ | ~1.9 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Ling 3.0 tiny yet to call it. We found 56 posts but people did not say much either way.
39 POSTS · 17 COMMENTS · LEMMY · GITHUB · REDDIT · X · BLUESKY
Other models the crowd has fully reviewed, starting with text models like this one.
inclusionAI/Ling-3.0-tiny · 8B A1.3B MoE· Hugging Face
Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!
AntLing’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, and WSM-merged stages.
🧵 We’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, an
ling 3.0 flash/tiny base models
Ling-3.0 (BailingMoE3) lands in llama.cpp mainline - Quick benchmarks on Intel Arc B580
Local AI News You Missed - August 2026
Agentic harness for small models
InclusionAI just open-sourced the Ling-3.0-tiny and Ling-3.0-flash Base Models: six checkpoints from pretraining to fina
I mapped the labs' free tiers in the article. These platforms made free access a daily habit, not a one-time trial: > To
noctrex/Ling-3.0-tiny-MXFP4_MOE-GGUF · Hugging Face
Alibaba Ant's Ling-3.0-tiny is now available as open-weight: BF16: huggingface.co/inclusionAI/... FP8: huggingface.co/i