🌳 1-bit Bonsai 27B
Full 27B-class reasoning with binary weights — 3.9 GB deployed (14.2× smaller than FP16), ~90% of FP16 intelligence retained, 262K-token context, at a true 1.125 bits/weight. This Space runs the Q1_0 GGUF build with the PrismML llama.cpp fork; the original model is prism-ml/Bonsai-27B-mlx-1bit.
⏳ Loading…
MultimodalTextbox
0 1.5
128 4096
Learn more: Whitepaper · Ternary Bonsai 27B · Prism ML · Discord
Note: this Space runs on free CPU hardware, so replies stream slowly — patience appreciated. It also sleeps when idle; the first message after a wake-up may take a few minutes while the model reloads.