LOADING DATASHEET
LOADING DATASHEET
by OpenAI
RANKED #111 OF 237 ASSISTANTS · OVERALL #189 OF 6,560 · VIBE SCORE 5.9 · 60 VOICES
Ling 3.0 Tiny is a compact 7.9B-parameter Mixture-of-Experts model with 1.3B parameters active per token. It is designed for responsive agents, reliable instruction following, natural multi-turn conversations, long-context work, and functio
14 mentions
0 mentions
26 weeks · 174 voices
QUIETLY DEGRADING
Recent vibe and chatter are both sliding.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Ling 3.0 Tiny (free) once and it's yours. No subscription, no rate limits, works offline.
7.9B parameters (Mixture of Experts, 1.3B active), a Q4 file is about 4.8 GB
NVIDIA GeForce RTX 3060 12GB
12 GB VRAM
~$220 USEDQ4_K_M, ~50 tok/s, 8k context
Apple Mac mini M4 (16GB RAM)
16 GB UNIFIED MEMORY
~$799 NEWFP8 or Q8_0, ~85 tok/s, 16k context
NVIDIA GeForce RTX 4090 24GB
24 GB VRAM
~$1600 USEDBF16 (full precision), ~120 tok/s, 131k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Ling 3.0 Tiny (free) takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3060 12GB | SWEET SPOTApple Mac mini M4 (16GB RAM) | FULL POWERNVIDIA GeForce RTX 4090 24GB |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~10 S★★★★★ | ~6 S★★★★★ | ~4 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~30 S★★★★★ | ~18 S★★★★★ | ~13 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~2.7 MIN★★★★★ | ~1.6 MIN★★★★★ | ~67 S★★★★★ |
| Build a backendan API with routes, storage and tests | ~8.3 MIN★★★★★ | ~4.9 MIN★★★★★ | ~3.5 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~5 MIN★★★★★ | ~2.9 MIN★★★★★ | ~2.1 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Ling 3.0 Tiny (free) yet to call it. We found 60 posts but people did not say much either way.
41 POSTS · 19 COMMENTS · REDDIT · X · LEMMY · GITHUB · BLUESKY
Other models the crowd has fully reviewed, starting with text models like this one.
inclusionAI/Ling-3.0-tiny · 8B A1.3B MoE· Hugging Face
Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!
AntLing’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, and WSM-merged stages.
🧵 We’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, an
ling 3.0 flash/tiny base models
Ling-3.0 (BailingMoE3) lands in llama.cpp mainline - Quick benchmarks on Intel Arc B580
Ling-3.0-tiny is showing where AI is heading: more capability without demanding more compute. Strong reasoning, agentic
Local AI News You Missed - August 2026
Agentic harness for small models
InclusionAI just open-sourced the Ling-3.0-tiny and Ling-3.0-flash Base Models: six checkpoints from pretraining to fina
I mapped the labs' free tiers in the article. These platforms made free access a daily habit, not a one-time trial: > To
noctrex/Ling-3.0-tiny-MXFP4_MOE-GGUF · Hugging Face