LOADING DATASHEET
LOADING DATASHEET
by Nvidia
RANKED #197 OF 337 ASSISTANTS · OVERALL #312 OF 528 RANKED · VIBE SCORE 5.6 · 44 VOICES
COPY BADGE pastes a README snippet with the rank SVG.
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
4 mentions
2 mentions
26 weeks · 66 voices
HONEYMOON FADING
Early praise is cooling off in recent weeks.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Nvidia Nemotron 3 Super 120B once and it's yours. No subscription, no rate limits, works offline.
120B total parameters (12B active MoE), a Q4 / NVFP4 quantized model requires around 70 to 75 GB memory
Apple Mac Studio M2 Max (96 GB RAM)
96 GB UNIFIED MEMORY
~$2400 USEDQ4 / 4-bit quant, ~16 tok/s, 8k context
Dual NVIDIA RTX 3090 (2x 24 GB) system with 128 GB system RA
48 GB VRAM + 128 GB RAM
~$1900 USEDQ4 / 4-bit quant with GPU offloading and system RAM, ~22 tok/s, 16k context
Apple Mac Studio M2 Ultra (192 GB RAM)
192 GB UNIFIED MEMORY
~$4800 USEDQ8 / 8-bit quant, full context window, ~35 tok/s
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Nvidia Nemotron 3 Super 120B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMApple Mac Studio M2 Max (96 GB RAM) | SWEET SPOTDual NVIDIA RTX 3090 (2x 24 GB) system with 128 GB system RA | FULL POWERApple Mac Studio M2 Ultra (192 GB RAM) |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~31 S★★★★★ | ~23 S★★★★★ | ~14 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~1.6 MIN★★★★★ | ~68 S★★★★★ | ~43 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~8.3 MIN★★★★★ | ~6.1 MIN★★★★★ | ~3.8 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~26 MIN★★★★★ | ~19 MIN★★★★★ | ~12 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~16 MIN★★★★★ | ~11 MIN★★★★★ | ~7.1 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Nvidia Nemotron 3 Super 120B yet to call it. We found 44 posts but people did not say much either way.
43 POSTS · 1 COMMENTS · REDDIT · GITHUB · BLUESKY
Other models the crowd has fully reviewed, starting with text models like this one.
NVIDIA releases Nemotron 3 Super!
My 10-Node Agentic RAG Architecture: Combining LangGraph, Cohere, Pinecone & MCP for Dense Legal Parsing
Free Coding Agent with NVIDIA Nemotron (Open Source)
I built a small proxy that lets Claude Desktop / Claude Code run on local models and NVIDIA's free API, sharing it in case it's useful
Phase 4: deterministic contract selector (4A chain observation + 4B selection)
#848: discovery живых model id NVIDIA NIM/Ollama Cloud/OpenRouter
fix(models): retire Gemma 4, move cheap tier to GLM 4.7 Flash + Qwen 3.5
Stop pre-throttling paid OpenRouter and Gemini keys
feat(triagem): OpenRouter no lugar do Gemini e preenchimento automatico do agente
fix(pricing): correct DeepSeek V4 rates, add ten IU gateway models
fix(dream): la nuit repart sur des primaires vivants — fin des 410 llama
fix(llm): replace retired Llama defaults; run gpt-oss at low reasoning effort