LOADING DATASHEET
LOADING DATASHEET
by BigCode
RANKED #8 OF 12 CODE · OVERALL #262 OF 6,567 · VIBE SCORE 5.7 · 36 VOICES
starcoder on Hugging Face (text generation). 9,977 downloads. Open weights for local or hosted use.
6 mentions
2 mentions
26 weeks · 10 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Starcoder once and it's yours. No subscription, no rate limits, works offline.
15.5B parameters, a Q4_K_M GGUF is about 9.5 GB (requires ~12-14 GB VRAM with context)
NVIDIA GeForce RTX 3060 12GB
12 GB VRAM
~$220 USEDQ4_K_M, ~22 tok/s, 4k context offloaded
NVIDIA GeForce RTX 4060 Ti 16GB
16 GB VRAM
~$430 NEWQ5_K_M or Q8, ~35 tok/s, full 8k context fully offloaded
NVIDIA GeForce RTX 4090 24GB
24 GB VRAM
~$1,750 USEDFP16 or Q8_0, ~75 tok/s, full 8k context with high throughput
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Starcoder takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3060 12GB | SWEET SPOTNVIDIA GeForce RTX 4060 Ti 16GB | FULL POWERNVIDIA GeForce RTX 4090 24GB |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~23 S | ~14 S | ~7 S |
| Summarize a documenta long report boiled down to the points that matter | ~68 S | ~43 S | ~20 S |
| Build a websitea small landing page, markup and styles together | ~6.1 MIN★★★★★ | ~3.8 MIN★★★★★ | ~1.8 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~19 MIN★★★★★ | ~12 MIN★★★★★ | ~5.6 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~11 MIN★★★★★ | ~7.1 MIN★★★★★ | ~3.3 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People are mostly happy with Starcoder. Biggest praise: help with code.
29 POSTS · 7 COMMENTS · LEMMY · REDDIT · BLOG · GITHUB
5 thumbs up · 1 thumbs down
“StarCoder 15b open-source code model beats Codex and Replit.”
“feat(gguf): llama-bpe verified via a mirror; starcoder and mpt refused for cause.”
Other models the crowd has fully reviewed, starting with code models like this one.
This week in AI - all the Major AI developments in a nutshell
StarCoder 15b open-source code model beats Codex and Replit
StarCoder-15B reaches 40.8% on HumanEval benchmark, beating the 30x bigger PaLM.
Nuggt: A LLM Agent the runs on WizardCoder-15B (4-bit quantised). It's time to democratise LLM Agents
New Starcoder 1B, 3B, and 7B models, Each model demonstrates the strongest performance for its size across various programming languages
What coding llm is the best?
Refact LLM: New 1.6B Code model with 32% HumanEval, SOTA for the size
Accelerate StarCoder with 🤗 Optimum Intel on Xeon: Q8/Q4 and Speculative Decoding
Creating a Coding Assistant with StarCoder
StarCoder: A State-of-the-Art LLM for Code
Make GitHub Copilot work with DeepSeek-Coder, Sonnet 3.5, Llama3, and Any Hosted or Ollama LLMs
Comment in r/singularity