My foray into local ai. Two BC-250 ex mining apus running Qwen3.6-35B-A3B Q4_K_M at 60 tok/s with 64k context
2.85T2 sourcer/LocalLLaMA
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1wou3gr/my_foray_into_local_ai_two_bc250_ex_mining_apus/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryA user reports running Qwen3-6-35B-A3B at Q4_K_M quantization on two BC-250 ex-mining APUs, achieving 60 tokens per second with a 64k context window as a local AI inference setup.
Why it mattersFirsthand benchmark of an unconventional ex-mining APU pairing for local LLM inference, with specific model, quantization, speed, and context numbers others can compare against.
Cited by
No citations on record.
