Ternary Bonsai 2 uses ternary weights to compress a Qwen3.8-27B architecture to less than 6GB, achieving a 9x reduction in size compared to FP16. The model retains 98.2% intelligence and is capable of running locally in-browser via WebGPU.
HOW THIS AFFECTS YOU
●
builderYou can deploy 27B parameter capabilities directly in user browsers using WebGPU without heavy backend infra.
●
designerYou can design highly responsive, privacy-first local AI interfaces that require zero latency from remote servers.