LOADING DATASHEET
LOADING DATASHEET
by Stepfun
RANKED #208 OF 311 ASSISTANTS · OVERALL #325 OF 6,952 · VIBE SCORE 5.6 · 29 VOICES
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters per token.
9 mentions
2 mentions
26 weeks · 40 voices
AGING WELL
The crowd is warmer now than it was at the start.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Step 3.7 Flash once and it's yours. No subscription, no rate limits, works offline.
198B sparse MoE (11B active), a Q4 file is about 122 GB (or ~82 GB in ROCmFPX Q3)
AMD Ryzen AI Max+ 395 laptop (128 GB unified LPDDR5X)
128 GB UNIFIED RAM
~$2,200 NEWROCmFPX Q3 or tight Q3_K_M quant with up to 32k context
Apple Mac Studio (M2/M3 Ultra, 192 GB Unified Memory)
192 GB UNIFIED MEMORY
~$5,500 USEDQ4_K_M quant with full native multimodal processing and long context
Dual NVIDIA RTX 6000 Ada workstation (96 GB total VRAM)
96 GB VRAM + 256 GB SYSTEM RAM
~$14,000 USEDNVFP4 or FP8 quantized checkpoint with full 256k context and low latency
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Step 3.7 Flash takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMAMD Ryzen AI Max+ 395 laptop (128 GB unified LPDDR5X) | SWEET SPOTApple Mac Studio (M2/M3 Ultra, 192 GB Unified Memory) | FULL POWERDual NVIDIA RTX 6000 Ada workstation (96 GB total VRAM) |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~36 S★★★★★ | ~21 S★★★★★ | ~11 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~1.8 MIN★★★★★ | ~63 S★★★★★ | ~33 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~9.5 MIN★★★★★ | ~5.6 MIN★★★★★ | ~3 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~30 MIN★★★★★ | ~17 MIN★★★★★ | ~9.3 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~18 MIN★★★★★ | ~10 MIN★★★★★ | ~5.6 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Step 3.7 Flash yet to call it. We found 29 posts but people did not say much either way.
21 POSTS · 8 COMMENTS · X · DEV FORUMS · REDDIT · GITHUB · BLOG
Other models the crowd has fully reviewed, starting with multimodal models like this one.
Step 3.7 Flash IQ4_XS GGUF with preserve_thinking
Comment in r/LocalLLaMA
Using local models with Hermes vs Claude code
Step 3.7 Flash quants rolling out today
We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?
Is LM Arena over?
Step-3.7-Flash Unsloth GGUF KLD Benchmarks
A trend we continue to see in open model releases is that the ecosystem is becoming more diverse, with an increasing number of organizations releasing a wide range of models. A year ago, open artifacts and the open model landscape more broadly were dominated…
top 10 free ai api providers i found this week. all in one place. 10 providers. $0 not clickbait. not engagement farming
Big week for open AI, with 25+ notable open-weight drops across every modality (from Victor M on 𝕏)
New release Stepfun 3.7 flash vs Deepseek V4 flash
Benchmarked the 128GB M5 Max as an Amateur - Need Feedback