parallelquant
July 14, 2026 · MarkTechPost

PrismML releases 1-bit and ternary quantized Qwen3.6-27B builds

PrismML released Bonsai 27B, a low-bit quantized version of Qwen3.6-27B rather than a new pretrained model. A ternary variant uses 1.71 bits per weight and fits in about 5.9GB, while a smaller 1-bit binary variant is also available. Both are released under the Apache 2.0 license and can run on laptops and phones.

Why it matters: Ultra-low-bit quantization lets large language models run on consumer hardware without retraining.

Related updates