LOADING DATASHEET
LOADING DATASHEET
by Unknown
RANKED #13 OF 17 CODE · OVERALL #410 OF 528 RANKED · VIBE SCORE 5.3 · 123 VOICES
COPY BADGE pastes a README snippet with the rank SVG.
Codellama available through cloud APIs.
11 mentions
12 mentions
26 weeks · 44 voices
HOLDING UP
Recent vibe is steady. No clear fade in the crowd.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Codellama once and it's yours. No subscription, no rate limits, works offline.
7B to 70B parameters; popular 7B Q4 is about 4.5 GB and 34B Q4 is about 20 GB
NVIDIA GeForce RTX 3060 12GB
12 GB VRAM
~$260 USEDCodeLlama 7B and 13B at Q4_K_M, ~35 tok/s, 8k context
NVIDIA GeForce RTX 4070 Ti Super 16GB
16 GB VRAM
~$750 NEWCodeLlama 34B at Q3_K_S or 13B at Q8_0, ~45 tok/s, 16k context
Apple Mac Studio M2 Ultra (64GB RAM)
64 GB UNIFIED MEMORY
~$3,400 USEDCodeLlama 70B at Q4_K_M or 34B at FP16, ~22 tok/s, 100k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Codellama takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3060 12GB | SWEET SPOTNVIDIA GeForce RTX 4070 Ti Super 16GB | FULL POWERApple Mac Studio M2 Ultra (64GB RAM) |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~14 S★★★★★ | ~11 S★★★★★ | ~23 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~43 S★★★★★ | ~33 S★★★★★ | ~68 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~3.8 MIN★★★★★ | ~3 MIN★★★★★ | ~6.1 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~12 MIN★★★★★ | ~9.3 MIN★★★★★ | ~19 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~7.1 MIN★★★★★ | ~5.6 MIN★★★★★ | ~11 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People are mostly happy with Codellama. Biggest praise: help with code.
95 POSTS · 28 COMMENTS · REDDIT · LEMMY · GITHUB · X
4 thumbs up · 0 thumbs down
“A new fine-tuned CodeLlama model called Phind beats GPT-4 at coding, 5x faster, and 16k context size.”
Other models the crowd has fully reviewed, starting with code models like this one.
JARVIS Local AGENT
I am terrified
This week in AI - all the Major AI developments in a nutshell
I created a free, open source Web extension to run Ollama
LLM finetuning 2-30X faster, use 60% less memory through OpenAI's Triton and mathematical tricks
The State of AI Engineering: notes from the first-ever AI Engineer Summit
Mistral 7B AI Model Released Under Apache 2.0 License
I built a social network where 6 Ollama agents debate each other autonomously — Mistral vs Llama 3.1 vs CodeLlama
Finetune LLMs 2-5x faster, use 60% less memory by using OpenAI's Triton language.
The complete tour of octogen, an opensource code interpreter powered by gpt4 and codellama
Sam Altman says ChatGPT should be 'much less lazy now'
What VS Code Extension works best with Ollama models?