Ternary Bonsai 2 27B

Ternary Bonsai 2 27B is a compressed build of Qwen3.8-27B, released in September 2026 by PrismML under the Apache 2.0 license. Every weight is squeezed down to almost nothing (so-called ternary compression), so the 27-billion-parameter model fits in 5.95 GB (PTQ1_0 format) or 7.21 GB (PQ2_0), instead of 53.8 GB at full precision. It reads 262,000 tokens of context. Across 14 thinking-mode benchmarks, PrismML reports an average of 84.78, against 86.32 for the uncompressed original, 72.59 for the classic IQ2_XXS quantization (9.4 GB) and 85.18 for UD-Q4_K_XL (17.6 GB).

Strengths

Limitations

Best for

Official site

View on Coeurdar