LOADING DATASHEET
LOADING DATASHEET
by Google
RANKED #22 OF 267 ASSISTANTS · OVERALL #36 OF 6,566 · VIBE SCORE 6.9 · 2,170 VOICES
A fast text model for people who need assistance with coding and writing tasks.
288 mentions
147 mentions
26 weeks · 2,888 voices
QUIETLY DEGRADING
Recent vibe and chatter are both sliding.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Gemma 4 26B A4B once and it's yours. No subscription, no rate limits, works offline.
25.2B parameters (MoE, 3.8B active), a Q4 file is about 14 to 17 GB
NVIDIA GeForce RTX 4060 Ti 16GB
16 GB VRAM
~$420 NEWQ4 quant (e.g., Q4_K_M or QAT Q4_0), ~30 tok/s, 8k context
NVIDIA GeForce RTX 3090 24GB
24 GB VRAM
~$700 USEDQ8 or Q5 quant, ~100 tok/s, 32k context
Apple Mac Studio M2 Max (64GB)
64 GB UNIFIED MEMORY
~$1700 USEDUnquantized BF16 or Q8, ~35 tok/s, full 128k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Gemma 4 26B A4B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 4060 Ti 16GB | SWEET SPOTNVIDIA GeForce RTX 3090 24GB | FULL POWERApple Mac Studio M2 Max (64GB) |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~17 S★★★★★ | ~5 S★★★★★ | ~14 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~50 S★★★★★ | ~15 S★★★★★ | ~43 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~4.4 MIN★★★★★ | ~80 S★★★★★ | ~3.8 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~14 MIN★★★★★ | ~4.2 MIN★★★★★ | ~12 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~8.3 MIN★★★★★ | ~2.5 MIN★★★★★ | ~7.1 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Users love Gemma 4 26B A4B for its speedy responses and strong help with coding and writing. However, some find it expensive and complain about occasional hallucinations, refusal lectures, and service errors.
1560 POSTS · 610 COMMENTS · REDDIT · BLUESKY · HACKER NEWS · X · GITHUB · LEMMY · DEV FORUMS · DEV.TO · BLOG
22 thumbs up · 2 thumbs down
13 thumbs up · 6 thumbs down
0 thumbs up · 19 thumbs down
9 thumbs up · 8 thumbs down
3 thumbs up · 12 thumbs down
5 thumbs up · 5 thumbs down
5 thumbs up · 4 thumbs down
1 thumbs up · 3 thumbs down
Checked picks first: tags say why you would switch, stars say how fully each one stands in for Gemma 4 26B A4B.
Quick tip: download the open models you care about from Hugging Face as soon as you can. You never know what the next fe
Running a 31B model locally made me realize how insane LLM infra actually is
Google releases new Gemma 4 QAT models!
Google is updating Gemma 4's chat templates, bringing major fixes to tool calling and reducing "laziness", and enabling Flash Attention 4 on Hopper GPUs, plus an interactive guide on how to work with and improve its vision!
Introducing Unsloth for AMD
Gemma 4 12B is out now!
Local AI News You Missed - May 2026
Google releases Gemma 4 models.
Google Gemma 4 MTP out now!
Google just dropped Gemma 4 12B on your laptop!!
Muse Glimmer ACTUALLY fits on a single RTX 3090
1 Day in and I feel okay saying Muse-Glimmer-30B finally beats 3.6-27B for the size in some use-cases