LOADING DATASHEET
LOADING DATASHEET
by Prism Ml
JUST LANDED · UPDATING HOURLY
RANKED #131 OF 309 ASSISTANTS · OVERALL #212 OF 6,980 · VIBE SCORE 6.0 · 82 VOICES
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks...
5 mentions
0 mentions
26 weeks · 4,021 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Ternary Bonsai 2 27B once and it's yours. No subscription, no rate limits, works offline.
27B parameters, a PTQ1_0 ternary GGUF is about 5.95 GB (or 8.6 GB for MLX with vision)
NVIDIA GeForce RTX 3060 12GB
12 GB VRAM
~$220 USEDPTQ1_0 1.75-bit pack, ~34 tok/s, 8k context
NVIDIA GeForce RTX 4070 Ti Super 16GB
16 GB VRAM
~$750 NEWPQ2_0 2.13-bit pack, ~55 tok/s, 64k context
Apple MacBook Pro M3 Max 64GB
64 GB UNIFIED RAM
~$2,800 USEDMLX 2-bit pack with full vision tower, ~47 tok/s, 262k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Ternary Bonsai 2 27B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3060 12GB | SWEET SPOTNVIDIA GeForce RTX 4070 Ti Super 16GB | FULL POWERApple MacBook Pro M3 Max 64GB |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~15 S | ~9 S | ~11 S |
| Summarize a documenta long report boiled down to the points that matter | ~44 S | ~27 S | ~32 S |
| Build a websitea small landing page, markup and styles together | ~3.9 MIN★★★★★ | ~2.4 MIN★★★★★ | ~2.8 MIN★★★★★ |
| Build a backendan API with routes, storage and tests | ~12 MIN★★★★★ | ~7.6 MIN★★★★★ | ~8.9 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~7.4 MIN★★★★★ | ~4.5 MIN★★★★★ | ~5.3 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
People generally like Ternary Bonsai 2 27B, with some gripes. Biggest praise: speed.
78 POSTS · 4 COMMENTS · GITHUB · HACKER NEWS · REDDIT · BLOG
2 thumbs up · 1 thumbs down
Other models the crowd has fully reviewed, starting with multimodal models like this one.
if you see this, it’s because you’re a real fan. AI News for 9/16/2026-9/17/2026. We checked 12 subreddits, 544 Twitters and no further Discords. AINews’ website lets you search all past issues. As a reminder, AINews is now a section of Latent Space . You can…
ICLR SUBMISSION 47647 how that possible? [D]
US chip fabs face massive 157,000 worker shortfall, mere 3% of US engineering grads enter chipmaking — despite six-figure salaries, US chip manufacturers are in dire need of engineers and technicians
US military had close call after using AI for false intelligence report, sources say
[ Removed by moderator ]
ggml : add PQ2_0 and PTQ1_0 ternary types; llama : apply Hadamard-folded weights
Bonsai 2 27B on an 8GB M2 Mac mini: 7.6 tok/s, every layer on GPU, no swap. The four flags that make it fit, and the one that kernel-panicked the machine
Comment in r/MachineLearning
Local models for smaller tasks, help?
chore(prices): sync OpenRouter prices: 6 models, 1 new [1 with gaps]
[Snyk] Security upgrade tap from 10.7.3 to 18.0.0
Performance: server-first product catalog