LOADING DATASHEET
LOADING DATASHEET
by MiniMax
RANKED #250 OF 327 ASSISTANTS · OVERALL #392 OF 7,100 · VIBE SCORE 5.4 · 25 VOICES
MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...
7 mentions
1 mentions
26 weeks · 20 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download MiniMax 01 once and it's yours. No subscription, no rate limits, works offline.
456B parameters (45.9B active MoE), a Q4 quantized file is about 260 GB
Custom workstation with 384 GB DDR5 RAM and 1x NVIDIA RTX 30
384 GB SYSTEM RAM PLUS 24 GB VRAM
~$2800 USEDToo large for single consumer GPUs; runs heavily offloaded to system RAM via CPU/GPU hybrid inference at Q3/Q4 with slow
Workstation with 8x NVIDIA RTX 3090 24 GB GPUs
192 GB VRAM PLUS 256 GB SYSTEM RAM
~$8500 USEDRuns 3-bit to 4-bit quantized (IQ3/Q4_K_M) across 8 GPUs in tensor parallelism at practical interactive speeds
Enterprise server with 8x NVIDIA H100 80 GB SXM
640 GB VRAM
~$240000 NEWFull FP8/BF16 precision inference with million-token context processing at high throughput
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long MiniMax 01 takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMCustom workstation with 384 GB DDR5 RAM and 1x NVIDIA RTX 30 | SWEET SPOTWorkstation with 8x NVIDIA RTX 3090 24 GB GPUs | FULL POWEREnterprise server with 8x NVIDIA H100 80 GB SXM |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~4.2 MIN★★★★★ | ~20 S★★★★★ | ~6 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~13 MIN★★★★★ | ~60 S★★★★★ | ~17 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~67 MIN★★★★★ | ~5.3 MIN★★★★★ | ~89 S★★★★★ |
| Build a backendan API with routes, storage and tests | ~3.5 HR★★★★★ | ~17 MIN★★★★★ | ~4.6 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~2.1 HR★★★★★ | ~10 MIN★★★★★ | ~2.8 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about MiniMax 01 yet to call it. We found 25 posts but people did not say much either way.
22 POSTS · 3 COMMENTS · REDDIT · GITHUB
Other models the crowd has fully reviewed, starting with multimodal models like this one.
MiniMax-01: Scaling Foundation Models with Lightning Attention. "our models match the performance of state-of-the-art models like GPT-4o and Claude-3.5-Sonnet while offering 20-32 times longer context window"
LLM Confabulation (Hallucination) Benchmark: DeepSeek R1, o1, o3-mini (medium reasoning effort), DeepSeek-V3, Gemini 2.0 Flash Thinking Exp 01-21, Qwen 2.5 Max, Microsoft Phi-4, Amazon Nova Pro, Mistral Small 3, MiniMax-Text-01 added
Anyone tried MiniMax-01 for coding? What's it like?
Why does livebench not benchmark MiniMax-01?
Has anyone tried using MiniMax-01 for long context roleplay?
Comment in r/SillyTavernAI
Has anyone tried MiniMax: MiniMax-01 with cline or roo-cline or any other Coding agent?
Model or workflow suggestions for large Json files?
Why does livebench not benchmark MiniMax-01?
Sorry guyd in laat post about minmax i gave the wrong link (minimax 02 is paid), im almost 100 percent certain 01 is opensource , tell me how you think it comapares to chatterbox
语音写稿换到 DeepSeek,等待从约 25 秒降到约 8 秒
Third-party API for DeepSeek R1 and MiniMax-Text-01 (flat monthly subscription)