LOADING DATASHEET
LOADING DATASHEET
by Google
RANKED #189 OF 279 ASSISTANTS · OVERALL #302 OF 6,613 · VIBE SCORE 5.6 · 28 VOICES
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
1 mentions
0 mentions
26 weeks · 16 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Gemma 3n 4B once and it's yours. No subscription, no rate limits, works offline.
4B parameters, a Q4 file is about 2.5 to 3.3 GB
Nvidia GeForce GTX 1650 (or 8 GB RAM laptop)
4 GB VRAM / 8 GB RAM
~$80 USEDQ4 quantization, ~15 tok/s, 4k context
Nvidia GeForce RTX 3060 12GB
12 GB VRAM
~$260 USEDQ8 / FP16 full offload, ~55 tok/s, 32k context
Nvidia GeForce RTX 4070 Ti Super 16GB
16 GB VRAM
~$750 NEWFP16 precision, ~85 tok/s, full 128k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Gemma 3n 4B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNvidia GeForce GTX 1650 (or 8 GB RAM laptop) | SWEET SPOTNvidia GeForce RTX 3060 12GB | FULL POWERNvidia GeForce RTX 4070 Ti Super 16GB |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~33 S | ~9 S | ~6 S |
| Summarize a documenta long report boiled down to the points that matter | ~1.7 MIN★★★★★ | ~27 S★★★★★ | ~18 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~8.9 MIN★★★★★ | ~2.4 MIN★★★★★ | ~1.6 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~28 MIN★★★★★ | ~7.6 MIN★★★★★ | ~4.9 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~17 MIN★★★★★ | ~4.5 MIN★★★★★ | ~2.9 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Gemma 3n 4B yet to call it. We found 28 posts but people did not say much either way.
27 POSTS · 1 COMMENTS · REDDIT · GITHUB
Other models the crowd has fully reviewed, starting with multimodal models like this one.
Gemma 3n Fine-tuning out now!
Gemma 3n $150,000 challenge
Unsloth GGUF + Model Updates: Gemma 3n fixed, MedGemma, Falcon, Orpheus, SmolLM, & more!
More Dynamic v2.0 GGUFs uploaded: Llama-4-Maverick, QwQ-32B, GLM-4-32B, Gemma-3-QAT, MAI-DS-R1 + more!
Google Gemma 3n Challenge ($150,000 in prizes) ends in 7 days! + New Gemma 3n notebooks
Gemma 3N Bug fixes + imatrix version
Choosing the right dataset format for dialogues
Gemma 3N E4B and Gemini 2.5 Flash Tested
VRAM Estimate Needed: Concurrent Gemma 3 4B Fine-tuning (GRPO) + vLLM Judge
Any context management features on the horizon?
Found this list from an api call. Is there any unreleased models/suprises?
Issue with finetuning Gemma 3 with "train_on_responses_only"