Ternary Bonsai 2 (27B) just released on Hugging Face. At <6GB in size, it can even run locally in-browser on WebGPU.
2.55T2 sourcer/LocalLLaMA
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1wj6c4l/ternary_bonsai_2_27b_just_released_on_hugging/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryTernary Bonsai 2, a 27B-parameter model using ternary quantization, was released on Hugging Face. The checkpoint is under 6GB, reportedly small enough to run locally in a browser via WebGPU.
Why it mattersA 27B model fitting under 6GB and running in-browser is a concrete data point on where local inference is heading — useful for anyone tracking self-hosted LLM setups.
Cited by
No citations on record.
