Speed on Bonsai 2 27b is interesting, especially in relation to the the MoE…
This is a dev post classified by Jev as Other (an opinion), kept by the Dev Radar because it carries real work, not commentary.
Speed on Bonsai 2 27b is interesting, especially in relation to the the MoE Qwen3.6-35b-a3b and its anticipated 3.8 version There is a size/speed tradeoff If you have 64 gb of memory for models, then it will probably be better to use the future Qwen 3.8 a3b because it will be faster, around 2x if it uses the same architecture But if you prioritize quality over speed TODAY, then Bonsai 2 27b is a better alternative because it is reportedly a high fidelity compression that would outperform a3b But if you have 32 to 16gb of memory, then rejoice! Your local model just got a huge bump in intell
Posted by Onur Solmaz (10k followers) 1 days ago · 81 likes · 12.8k views · view the original post on X. Kept by the Dev Radar as Other. Tools mentioned: Gist, ternary-bonsai-2-webgpu-kernels.
More dev work like this
- You can run locally Uncensored Ternary Bonsai 2-27B Same 5.9GB file. — @0x0SojalSec
- dear rtx 3060 owners, and every 12gb card behind it. bonsai 2 27b dense went from 26 to… — @sudoingX
- Nice: banked codex reset for all of us. — @kimmonismus
- Convolutional Neural Networks. #BigData #Analytics #DataScience #AI #MachineLearning… — @gp_pulipaka
- Reinforcement Learning and Optimal Control! #BigData #Analytics #DataScience #AI… — @gp_pulipaka
- The #Mathematical #Programming - #FPGA. #BigData #Analytics #DataScience #AI… — @gp_pulipaka
- A Constraint Based Approach, ML. #BigData #Analytics #DataScience #AI #MachineLearning… — @gp_pulipaka
- FBX対応も問題無さそう — @koguGameDev
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 14.4k posts from 4.7k X accounts over the last 21 days, 1.6k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 21:10 UTC. Full method.