Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size
1.95T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1vjanps/kimi_k3_unsloth_iq2xxs_from_711gb_down_to_478gb/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryCommunity post reports the Kimi K3 model quantized with Unsloth's IQ2-XXS method, shrinking the files from 711GB to 478GB by removing multilingual capability to reduce size.
Why it mattersConcrete numbers for an extreme quantization tradeoff on a frontier-scale model, useful for anyone fitting very large weights onto local hardware.
Cited by
No citations on record.
