How PrismML’s Bonsai Compresses a 27B Model to 3.9 GB for Browser and Phone Deployment
PrismML’s Bonsai compresses the 27B Qwen3.6 model to a 3.9 GB 1‑bit version that runs on iPhone 17 Pro and in browsers via WebGPU, achieving competitive benchmark scores, high "intelligent density", and enabling zero‑cost, privacy‑preserving agent applications on edge devices.
