27B on a Phone: The Density Numbers Behind PrismML’s Bonsai Breakthrough
PrismML Bonsai 27B uses ternary quantization to compress a 27B model to 5.9GB, enabling phone deployment. 10x intelligence density versus full-precision. Benchmark gaps analyzed.