LOADING DATASHEET
LOADING DATASHEET
by Microsoft
RANKED #172 OF 271 ASSISTANTS · OVERALL #274 OF 6,567 · VIBE SCORE 5.6 · 101 VOICES
[Microsoft Research](/microsoft) Phi-4 is designed to perform well in complex reasoning tasks and can operate efficiently in situations with limited memory or where quick responses are needed. At 14 billion...
18 mentions
8 mentions
26 weeks · 56 voices
HONEYMOON FADING
Early praise is cooling off in recent weeks.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Phi 4 once and it's yours. No subscription, no rate limits, works offline.
14B parameters, a Q4 file is about 9 GB
NVIDIA GeForce RTX 3060 12GB
12 GB VRAM
~$220 USEDQ4_K_M, ~25 tok/s, 16k context
NVIDIA GeForce RTX 4060 Ti 16GB
16 GB VRAM
~$380 USEDQ8_0, ~35 tok/s, 16k context
NVIDIA GeForce RTX 4090 24GB
24 GB VRAM
~$2,250 USEDQ8_0 or FP16 with partial offload, ~60 tok/s, 16k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Phi 4 takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3060 12GB | SWEET SPOTNVIDIA GeForce RTX 4060 Ti 16GB | FULL POWERNVIDIA GeForce RTX 4090 24GB |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~20 S★★★★★ | ~14 S★★★★★ | ~8 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~60 S★★★★★ | ~43 S★★★★★ | ~25 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~5.3 MIN★★★★★ | ~3.8 MIN★★★★★ | ~2.2 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~17 MIN★★★★★ | ~12 MIN★★★★★ | ~6.9 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~10 MIN★★★★★ | ~7.1 MIN★★★★★ | ~4.2 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Phi 4 yet to call it. We found 101 posts but people did not say much either way.
74 POSTS · 27 COMMENTS · REDDIT · OTHER FORUMS · HACKER NEWS · GITHUB · X · LEMMY
Checked picks first: tags say why you would switch, stars say how fully each one stands in for Phi 4.
You can now train your own o3-mini model on your local device!
I fixed 4 bugs in Microsoft's open-source Phi-4 model
[P] How I found & fixed 4 bugs in Microsoft's Phi-4 model
[R] LLMs are Locally Linear Mappings: Qwen 3, Gemma 3 and Llama 3 can be converted to exactly equivalent locally linear systems for interpretability
Train your own Reasoning model like DeepSeek-R1 locally (7GB VRAM min.)
You can now train your own DeepSeek-R1 model on your local device!
Is Microsoft-Phi dead?
You can now run 'Phi-4 Reasoning' models on your own local device! (20GB RAM min.)
PHI-4: Quantized Q4 vs Q8 on My Nvidia RTX 3060 12GB System
Train your own Reasoning model like DeepSeek-R1 locally (5GB VRAM min.)
I benchmarked 21 local LLMs on a MacBook Air M5 for code quality AND speed
Show HN: A new benchmark for testing LLMs for deterministic outputs