I’m calling this the Monstrosity. 5 ex mining BC-250 boards Qwen3-Coder-Next Q4 at 40 tok/s
2.40T2 sourcer/LocalLLaMA
r/LocalLLaMAoriginal source ↗
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1ws49si/im_calling_this_the_monstrosity_5_ex_mining_bc250/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryUser reports running Qwen3-Coder-Next at Q4 quantization across 5 repurposed BC-250 mining GPU boards, achieving 40 tokens/second inference.
Why it mattersUnusual build log showing ex-mining boards repurposed for local LLM inference, with concrete throughput numbers worth noting for anyone weighing similar hardware.
Cited by
No citations on record.
