LOADING DATASHEET
LOADING DATASHEET
by Moonshot
RANKED #47 OF 228 ASSISTANTS · OVERALL #79 OF 6,618 · VIBE SCORE 6.4 · 350 VOICES
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
62 mentions
21 mentions
26 weeks · 126 voices
HOLDING UP
Recent vibe is steady. No clear fade in the crowd.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Kimi K2 once and it's yours. No subscription, no rate limits, works offline.
1T total parameters (32B active MoE); a Q4 quant requires roughly 550 GB to 600 GB of memory.
Custom multi-GPU rig (e.g. 8x NVIDIA RTX 3090 24GB used) or
192 GB VRAM PLUS 512 GB SYSTEM RAM
~$7500 USEDModel is too large for standard consumer hardware. At heavy offloading/Q2/Q3 it runs very slowly at 1 to 4 tok/s.
Apple Mac Studio M2 Ultra cluster or multi-GPU workstation (
192 GB TO 512 GB UNIFIED VRAM
~$16000 BUILTINT4 or Q4 quant via KTransformers/vLLM, ~15 to 25 tok/s, 32k context
Dedicated enterprise node with 8x NVIDIA H100 80GB SXM
640 GB HBM3
~$280000 NEWNative FP8/BF16, full throughput at 60+ tok/s, full 256k context window
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Kimi K2 takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMCustom multi-GPU rig (e.g. 8x NVIDIA RTX 3090 24GB used) or | SWEET SPOTApple Mac Studio M2 Ultra cluster or multi-GPU workstation ( | FULL POWERDedicated enterprise node with 8x NVIDIA H100 80GB SXM |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~3.3 MIN★★★★★ | ~25 S★★★★★ | —★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~10 MIN★★★★★ | ~75 S★★★★★ | —★★★★★ |
| Build a websitea small landing page, markup and styles together | ~53 MIN★★★★★ | ~6.7 MIN★★★★★ | —★★★★★ |
| Build a backendan API with routes, storage and tests | ~2.8 HR★★★★★ | ~21 MIN★★★★★ | —★★★★★ |
| Build a gamea playable browser game in one file | ~1.7 HR★★★★★ | ~13 MIN★★★★★ | —★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. A DASH MEANS NO SPEED WAS RESEARCHED FOR THAT RIG. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People are mostly happy with Kimi K2. Biggest praise: writing. Biggest gripe: saying no too much.
292 POSTS · 58 COMMENTS · REDDIT · HACKER NEWS · X · GITHUB · LEMMY · BLUESKY · DEV.TO · BLOG
11 thumbs up · 1 thumbs down
5 thumbs up · 4 thumbs down
9 thumbs up · 0 thumbs down
4 thumbs up · 0 thumbs down
4 thumbs up · 0 thumbs down
2 thumbs up · 2 thumbs down
1 thumbs up · 2 thumbs down
“Both were first try as well, but Opus is so much more expensive :/.”
Checked picks first: tags say why you would switch, stars say how fully each one stands in for Kimi K2.
The chinese did it, KIMI K2 surpassed GPT-5.
Kimi K2 Thinking, a SOTA open-source trillion-parameter reasoning model
We are accelerating faster than people realise. Every week is overwhelming
Car Wash Test on 53 leading models: “I want to wash my car. The car wash is 50 meters away. Should I walk or drive?”
Indeed, Composer 2 is kimi k2
Kimi K2 Thinking is now available on Perplexity
INCREDIBLE STUFF INCOMING
Kimi K2 Thinking, A Chinese Open-Source Trillion-Parameter Thinking model, surpass Grok 4 and GPT-5 on HLE
Kimi K2: New SoTA non-reasoning model 1T parameters open-source and outperforms DeepSeek-v3.1 and GPT-4.1 by a large margin
Made the switch to DeepSeek and here are my thoughts as a long time Claude user (spoiler: it's great)
Kimi K2 is already irrelevant, and it's only been like 1 week. Qwen has updated Qwen-3-235B, and it outperforms K2 at less than 1/4th the size
Kimi k2 an open source just surpassed o3 in creative writing and eq bench !!