PrismML Releases Ternary Bonsai 2 27B: A 5.9 GB Apache 2.0 Model Retaining 98.2% of Qwen3.8 27B Performance
PrismML released Ternary Bonsai 2 27B, a 5.93 GB Apache 2.0 ternary-weight model retaining 98.2% of Qwen3.8 27B's benchmark performance.
PrismML released Ternary Bonsai 2 27B, a ternary-weight version of Qwen3.8 27B that shrinks the model from 53.80 GB FP16 to 5.93 GB under Apache 2.0 licensing. It scores 83.9 across 20 benchmarks versus 85.4 for the FP16 parent, and outperforms an IQ2_XXS quantization on AIME26 (95.83 vs 78.6) and LiveCodeBench v6 (90.07 vs 70.05). The 27.36B-parameter model supports text and images with a 262K-token context and runs on a 16 GB laptop or single 24 GB GPU via PrismML's llama.cpp fork or MLX runtime. Long-horizon agentic results retain only about 75% of full-precision scores, and all benchmarks are PrismML's own, not independently reproduced.