PrismML hopes its tiny LLM will change how we all use AI
PrismML released Bonsai 2 27B, compressing Alibaba's Qwen 27B to 5.9 GB with ternary weights while retaining 98% of benchmark performance.
PrismML, a Caltech-founded startup with a $22.25 million seed round from Khosla Ventures, Cerberus Capital, and Caltech, released Bonsai 2 27B, which compresses Qwen3.8 27B down to 5.9 GB, a 9-10x memory reduction small enough for a PC or high-end smartphone. The model uses ternary weights (+1, -1, 0) instead of 16-bit values and matches 98% of Qwen's aggregate benchmark scores, up from 95% for the original Bonsai, which has been downloaded over 11 million times. CEO Babak Hassibi said upcoming releases will target compressed models in the several-hundred-billion-parameter range, and the startup is rumored to be in talks with Apple.